Convert Scanned PDFs to Text

Job ID: 40383054

Budget: ₹600 – ₹1,500 INR

I have a batch of PDFs made up entirely of scanned images. The scans are of medium quality, so the text is readable but not crystal-clear; careful OCR work and manual proofreading will be required. All I need back is clean, plain UTF-8 text—no layout reconstruction, styling, or headers necessary.

Speed is important: this project is time-sensitive and I’m ready to start as soon as I select the freelancer. Tools such as Tesseract, ABBYY FineReader, or Adobe Acrobat’s OCR module are fine, provided the final output meets the accuracy standard below.

Deliverables
• A separate .txt file for each source PDF, named to match the original file
• Spelling and obvious recognition errors corrected
• Any unreadable characters clearly flagged with “[unreadable]” so I can review them quickly

Acceptance criteria
• 99% character-level accuracy across the set, verified by spot checks
• Delivery within the agreed deadline (please state how soon you can finish when you bid)

If this feels straightforward for you, let me know your timeframe and a quick note on the OCR workflow you prefer.