Mixed Image Text Transcription
Budget: $15 – $25 USD
I have a collection of scanned pages—letters, receipts, and forms—where handwritten notes sit beside printed paragraphs. I need every word transcribed with care for long-term archival use, so accuracy matters more than speed. Each image should become its own UTF-8 .txt file, named to match the original, with illegible portions marked “[unclear]”.
Please keep original line breaks, spelling, and punctuation. Cross-outs or overwritten text must be reproduced exactly, followed by a brief bracketed note. Feel free to lean on OCR tools such as Tesseract, Adobe Acrobat, or ABBYY, but every page must be proof-read manually before delivery.
Deliverables
• Plain-text file for each image
• Short log listing any “[unclear]” tags and their locations
I will send the images in 100-page batches and appreciate your estimated turnaround time per batch along with any questions you may have.
Please keep original line breaks, spelling, and punctuation. Cross-outs or overwritten text must be reproduced exactly, followed by a brief bracketed note. Feel free to lean on OCR tools such as Tesseract, Adobe Acrobat, or ABBYY, but every page must be proof-read manually before delivery.
Deliverables
• Plain-text file for each image
• Short log listing any “[unclear]” tags and their locations
I will send the images in 100-page batches and appreciate your estimated turnaround time per batch along with any questions you may have.
Related categories:
Data Processing
Data Entry
Proofreading
OCR
ABBYY FineReader
Adobe Acrobat
Text Recognition
Data Annotation