
Closed
Posted
Paid on delivery
I need a Docker package and a Compose YAML file for ocrmypdf-paddleocr. Requirements: - Base Image: Ubuntu - Include Libraries: Python, OCRmyPDF, PaddlePaddle, PaddleOCR, Pillow - Primary Function: Convert PDF to text-searchable PDFs - GPU acceleration with Nvidia GPU - Monitor a directory (to be specified in docker compose file), and automatically run whenever a new PDF is added. Ideal Skills: - Experience with Docker and Docker Compose - Familiarity with OCR tools and Python libraries - Proficiency in creating Dockerfiles and managing dependencies
Project ID: 40595219
128 proposals
Remote project
Active 1 day ago
Set your budget and timeframe
Get paid for your work
Outline your proposal
It's free to sign up and bid on jobs
128 freelancers are bidding on average $139 AUD for this job

I am well-versed in Docker and OCR technologies, ensuring an efficient and scalable solution for your OCRmyPDF-PaddleOCR needs on Ubuntu. By leveraging Python, OCRmyPDF, PaddlePaddle, PaddleOCR, and Pillow, GPU acceleration with Nvidia GPUs will be integrated seamlessly. The solution will automate PDF conversion to text-searchable PDFs upon detecting new PDFs in a specified directory. With a focus on usability and future enhancements, I will create a well-structured Docker package and Compose YAML file to meet your requirements effectively. My expertise in OCR tools, Python libraries, and GPU acceleration guarantees optimal performance for extensive PDF processing tasks. I am committed to delivering a solution that aligns with your current and future needs, fostering a long-term partnership based on efficiency and reliability. I am eager to collaborate on this project and discuss further details to surpass your expectations.
$225 AUD in 5 days
8.8
8.8

Hello, I understand you need a Docker setup for the OCR tool 'ocrmypdf-paddleocr' using Ubuntu as the base image. Here's how I can assist you: 1. I will create a Dockerfile starting with the Ubuntu image, ensuring lightweight and optimized configuration. 2. Necessary libraries like Python, OCRmyPDF, PaddlePaddle, PaddleOCR, and Pillow will be installed within the Docker environment to ensure smooth operation of OCR tasks. 3. I will write a Docker Compose YAML file to orchestrate the container setup, simplifying the deployment and management process for you. 4. Additionally, I will ensure the setup supports automated container updates and conforms to best DevOps practices. This structured approach will provide you with a robust, scalable solution for your OCR needs. Best Regards, Khorshed Alam, RS Software
$100 AUD in 3 days
9.5
9.5

Hi, I will deliver a Dockerfile (Ubuntu base) and Compose YAML that runs ocrmypdf with PaddleOCR, PaddlePaddle, and Pillow, with NVIDIA GPU passthrough configured. The watch directory will be volume mounted in Compose so new PDFs trigger conversion automatically. On a similar setup, I used inotifywait to detect new files and queue them, which handled bursts of uploads without duplicate processing. Questions: 1) Should the watcher process files recursively in subdirectories? 2) Do you need the output PDFs written to a separate directory or in place? Share your GPU model and CUDA version and I will confirm the correct PaddlePaddle base. Looking forward to discussing further. Best regards, Kamran
$90 AUD in 5 days
8.5
8.5

Hi, I’m a senior developer with 20+ years of experience, Top Rated on Freelancer.com with 1,500+ completed projects and a 5-star record, and I’ve delivered similar Docker setups before, so this is straightforward. I’ll prepare a Dockerfile for Ubuntu with Python, OCRmyPDF, PaddlePaddle, PaddleOCR, and Pillow, then build it with GPU support using Nvidia’s runtime. The Compose file will include a volume mount for the monitored directory, an OCR service triggered by file changes, and a cleanup step to remove temporary files. I’ll ensure the container restarts automatically if it crashes and includes a health check to verify the OCR process completes successfully. The setup will log errors for debugging and handle file permissions consistently. I can start immediately.
$30 AUD in 3 days
7.9
7.9

Hi, I specialize in turning business challenges into clean, reliable software — from automation and APIs to web applications, AI, and backend systems. What sets me apart: I focus on outcomes, not just deliverables. Every solution I build is designed to be maintainable, scalable, and aligned with your actual goals. Let's build something great together. Thanks, Deva
$200 AUD in 2 days
7.8
7.8

A GPU-accelerated ocrmypdf-paddleocr container that watches a folder and auto-converts new PDFs into searchable ones is a clean, well-defined build, and it's right in my wheelhouse. I'll write a Dockerfile on an Ubuntu base with CUDA support, installing Python, OCRmyPDF, GPU-enabled PaddlePaddle, PaddleOCR, and Pillow with pinned versions so builds stay reproducible. For the auto-processing I'll add a lightweight watcher (watchdog/inotify) that detects new PDFs in the mounted input directory and runs OCRmyPDF using PaddleOCR as the engine, writing the searchable output to a target folder. The Compose file will wire up the input/output volume mounts and enable NVIDIA GPU access via the runtime and deploy reservations, with configurable paths. I can start immediately.
$30 AUD in 1 day
7.6
7.6

Hi, The directory watch is the part most Docker setups get wrong here. I'd run a watcher (inotify or watchdog) as the container entrypoint so it triggers ocrmypdf on each new PDF, rather than a cron poll, and mount the input directory as a volume defined in the Compose file. For the GPU side, the image needs the CUDA base plus nvidia runtime in Compose so PaddlePaddle actually sees the card, otherwise it silently falls back to CPU. One question: is the host running the Nvidia Container Toolkit already, and which CUDA version? That decides the base image tag. I handle Docker, CI/CD and Python deploys daily as part of MangoCoders. Happy to start on a milestone so you only release once the container runs on your PDFs. Adil
$183.70 AUD in 7 days
7.5
7.5

Hi, I am interested to work on this project.I have good experience with Python Docker Automation Linux APIs Development Looking forward to an early and positive response. Regards, Shalu
$130 AUD in 5 days
6.9
6.9

Hi, I can package your OCR workflow in Docker using Ubuntu, OCRmyPDF, and PaddleOCR with Nvidia GPU support, then set up an inotify-based file watcher in the container to trigger conversions automatically. I’ve built similar GPU-accelerated Python pipelines before, though the first iteration needed a few tweaks to stabilize the CUDA dependency chain. I’d structure this as a multi-stage Dockerfile to keep the image lean, use Nvidia’s runtime for GPU access, and run a small Python daemon inside the container to monitor the directory without adding complexity outside Docker. The setup will handle the PDF-to-searchable-PDF conversion efficiently while remaining easy to update or deploy elsewhere. If the file watcher behaves as expected, the whole system should run unattended with minimal maintenance. I can start right now. Thanks, Denis
$100 AUD in 2 days
6.2
6.2

Hi, I can build a Dockerized OCR solution that meets your requirements, including an Ubuntu-based image, OCRmyPDF + PaddleOCR + PaddlePaddle, NVIDIA GPU acceleration, and automatic directory monitoring to process new PDFs into searchable PDFs. Here's my approach: Ubuntu-based Docker image with all required dependencies Docker Compose with configurable watch directory GPU support using NVIDIA Container Toolkit Automatic PDF monitoring & processing Clean, documented setup with an easy deployment guide A couple of quick questions: Which CUDA version is your NVIDIA environment running? Should processed PDFs overwrite the originals or be saved to a separate output folder? Ready to get started immediately. Regards, Ravi B.
$120 AUD in 7 days
6.2
6.2

Hello! We can package this OCR workflow into a Docker setup for your current tasks. 1. Which folder should the container monitor for new PDFs? 2. Do you need GPU-only processing or a fallback without GPU? — About us We are dZENcode – a full-cycle IT company for digital product development: from design and programming to integrations and post-release support. We build projects from scratch and also work on existing solutions that need further development, improvements, or technical support. You can find detailed information about our services and rates on our official website: https://dzencode.com. Please review it – after that, we can discuss the details and agree on the next step. ⚠️ After clarifying all details, we will define the scope, the suitable cooperation format – task-based, outsourcing, or outstaffing – and the final cost. Projects are guaranteed to reach release with us: • 10+ years providing IT services; • 90+ in-house specialists; • 250+ public reviews since 2015; • We support products under SLA after launch; • We work under NDA and a company contract!
$140 AUD in 7 days
6.6
6.6

Hello There! I’m Md. Toriqul Islam, and I’m excited to partner with you. I can dive into your project immediately. I have over 10 years of experience with Docker, Docker Compose, Python, Ubuntu, OCR automation, GPU-enabled containers, and Linux environments. I understand you need a Dockerized OCRmyPDF + PaddleOCR solution running on Ubuntu with NVIDIA GPU support, automatic folder monitoring, and Docker Compose configuration. I can deliver a clean, production-ready setup with all dependencies installed, automatic PDF processing, and clear documentation for deployment and configuration. I am skilled in Docker, Docker Compose, Python, OCRmyPDF, PaddleOCR, PaddlePaddle, NVIDIA CUDA, and Linux. I’m ready to start immediately and would be happy to discuss your environment and GPU requirements. Looking forward to hearing from you. Best regards, Md. Toriqul Islam
$100 AUD in 3 days
6.1
6.1

The main challenge here is packaging the OCR stack correctly, especially with PaddleOCR GPU support and automatic PDF processing inside a container. I’d create an Ubuntu-based image with the required Python dependencies, configure NVIDIA runtime support, and add a lightweight watcher process so new PDFs in the mounted directory are processed without manual runs. I’d also keep the Compose setup configurable so paths, GPU settings, and future OCR options can be changed without rebuilding the image. Do you already have a target NVIDIA CUDA version/environment, or should the container be built around a recommended compatible stack?
$50 AUD in 1 day
5.8
5.8

Drawing upon my decade of expertise in Linux System Administration and DevOps, I am a robust candidate to handle your Docker setup needs for the OCR tool. In parallel to your project requirements, I have considerable experience in creating Dockerfiles, managing dependencies, and even optimizing performance on different Linux distros — making Ubuntu-based images a comfortable terrain for me. My consistent work with prominent AI libraries like TensorFlow and PyTorch, combined with GPU integration using CUDA or ROCM, can ensure smooth sailing even with Nvidia GPUs working overtime for faster OCR processing.
$250 AUD in 7 days
6.2
6.2

I can package OCRmyPDF with PaddleOCR into a reproducible Ubuntu-based Docker setup that automatically converts incoming PDFs into searchable PDFs using NVIDIA GPU acceleration. The deliverables will include a production-ready Dockerfile, Docker Compose YAML, Python watcher service, dependency lock file, health check, logging, and setup documentation. The input, output, processed, and failed directories will be configurable through Compose volumes and environment variables. My two priorities would be reliable processing and maintainability. The watcher will wait until a new PDF has finished copying before starting OCR, prevent duplicate processing, preserve the original file, and move failed jobs into a separate directory with clear error logs. CUDA, PaddlePaddle, PaddleOCR, OCRmyPDF, Pillow, and system packages will be pinned to compatible versions so the image remains reproducible. I will also add GPU availability checks, configurable OCR language and quality settings, restart handling, and a sample PDF validation to confirm that the output contains a searchable text layer. A relevant project is Voiceup, where I built an automated pipeline for processing incoming call recordings, generating transcripts and insights, and exposing processing results through role-based dashboards.
$140 AUD in 7 days
5.8
5.8

Hi there, we are a team of AI/ML Full Stack Web and Mobile App Developers and we can do this project in no time. Thanks Ashish Kumar.
$140 AUD in 7 days
5.9
5.9

Hi, The challenging part of this project isn't installing OCRmyPDF or PaddleOCR—it's building a stable Docker environment with proper NVIDIA GPU support and an automated file-watching workflow. I can create a clean Ubuntu-based Docker image including Python, OCRmyPDF, PaddlePaddle, PaddleOCR, and Pillow, along with a Docker Compose configuration where the watched directory is configurable. The container will automatically detect new PDFs, process them into searchable PDFs, and keep the setup easy to maintain. I'll also ensure the project is well documented so you can update paths or rebuild the container without hassle.
$100 AUD in 3 days
5.4
5.4

The main challenge is binding GPU support for PaddleOCR inside Docker and setting up automatic directory monitoring. OCRmyPDF and PaddleOCR both need specific CUDA drivers that match your Nvidia card and host OS. I’ve built similar GPU-accelerated OCR images and automation flows before. I write Dockerfiles on Ubuntu base, install all the Python libraries (OCRmyPDF, PaddlePaddle, PaddleOCR, Pillow), and set up inotify in the container for directory watching. Docker Compose handles volume mapping and the watch path. Does your host already run nvidia-docker runtime and drivers, or do you need those added to the Compose stack as well? Pradeep
$140 AUD in 7 days
5.3
5.3

As a seasoned technology partner, my skill set and experience uniquely position me to tackle your Docker setup for ocrmypdf-paddleocr. Throughout my career, I've developed web, mobile, SaaS applications while also creating scalable custom software, which aligns with your project's needs. My core technical competencies - Python, Docker & Docker Compose - alongside familiarity with OCR tools enhance my ability to deliver a seamless package for you. In addition, my proficiency in managing dependencies will ensure efficient workflows within the Docker containers. I understand the vital role of base image selection and libraries inclusion in optimizing application performance. So rest assured your solution will be built on the right base image (Ubuntu), with critical libraries such as; PaddleOCR, OCRmyPDF, Python and Pillow included - all of which are necessary to achieve the primary function you're seeking: PDF conversion to text-searchable formats. Lastly, my knowledge and experience extend to dealing with GPU acceleration when working with Nvidia GPU. This is undoubtedly crucial for achieving increased efficiency and performance - a unique aspect I bring to the table for this project. All these skills, honed over years of practice in different industries and international contexts, make me an ideal fit for ensuring your docker setup for OCR tool exceeds expectations in every regard. Let's maximize the potential of your project together!
$100 AUD in 6 days
5.3
5.3

Hey there! I'm really pumped about this opportunity! I recently led a project with similar challenges and nailed it. Drawing from my experience in PHP, Python, Linux, Software Architecture, Ubuntu, Docker, DevOps, Automation, Docker Compose, Containerization, I’m ready to dive into your project. Please come over chat and discuss your requirement in a detailed way. Cheers, Vishal Maharaj
$250 AUD in 7 days
5.0
5.0

Perth, Australia
Member since Jul 21, 2026
₹1500-12500 INR
₹1500-12500 INR
$30-250 AUD
$30-250 USD
€250-750 EUR
₹1500-12500 INR
₹600-1500 INR
€30-250 EUR
₹1500-12500 INR
€30-250 EUR
₹1250-2500 INR / hour
₹750-1250 INR / hour
₹1500-12500 INR
₹12500-37500 INR
$30-250 USD
₹37500-75000 INR
min €36 EUR / hour
$10-30 USD
₹400-750 INR / hour
$2-8 AUD / hour