OCR & Print Four Scanned PDFs
Budget: $10 – $30 USD
I have an email that itself is a single PDF. Buried inside that file are four scanned-in PDFs. I need two things done for each of those four scans:
1. Extract every legible word into a clean plain-text file (simple .txt is perfect—no formatting needed).
2. Return a print-ready version of the scan so I can send it straight to a standard office printer without margin or resolution issues.
You may use the OCR tool of your choice—Tesseract, Adobe Acrobat, ABBYY, or any equivalent—as long as the output is accurate and the page layout of the printable version matches the original scan.
Deliverables:
• 4 text files, one per scanned PDF.
• 4 optimized PDFs ready to print.
That’s the full scope; no extra formatting or design work required. Let me know if you can turn this around quickly and what OCR accuracy you typically achieve.
1. Extract every legible word into a clean plain-text file (simple .txt is perfect—no formatting needed).
2. Return a print-ready version of the scan so I can send it straight to a standard office printer without margin or resolution issues.
You may use the OCR tool of your choice—Tesseract, Adobe Acrobat, ABBYY, or any equivalent—as long as the output is accurate and the page layout of the printable version matches the original scan.
Deliverables:
• 4 text files, one per scanned PDF.
• 4 optimized PDFs ready to print.
That’s the full scope; no extra formatting or design work required. Let me know if you can turn this around quickly and what OCR accuracy you typically achieve.
Related categories:
PDF
Adobe InDesign
Word
OCR
Data Extraction
ABBYY FineReader
Adobe Acrobat
Text Recognition