Data entry of books
Budget: $250 – $750 USD
I have a set of complete books already digitised as PDFs and I need their text pulled out accurately and cleanly. Only the text matters—no images, tables, or formatting beyond the natural chapter and paragraph breaks that appear in the source.
Your task is simple but detail-oriented: open each PDF, extract every word, and place it into a separate, editable document without introducing typos or dropping content. Consistency counts; headings, sub-headings, and page breaks should flow naturally so the final file reads like the original book, just in plain text.
Deliverables
• One .docx or .txt file per PDF, named to match the source
• Text proof-read for obvious OCR errors and strange line breaks
If you already work with reliable OCR tools such as Adobe Acrobat, ABBYY FineReader, or Tesseract and know how to spot-check the output quickly, you’ll finish this smoothly. Let me know how many pages you can handle per day and when you can start; I’m ready to share the first batch right away.
Your task is simple but detail-oriented: open each PDF, extract every word, and place it into a separate, editable document without introducing typos or dropping content. Consistency counts; headings, sub-headings, and page breaks should flow naturally so the final file reads like the original book, just in plain text.
Deliverables
• One .docx or .txt file per PDF, named to match the source
• Text proof-read for obvious OCR errors and strange line breaks
If you already work with reliable OCR tools such as Adobe Acrobat, ABBYY FineReader, or Tesseract and know how to spot-check the output quickly, you’ll finish this smoothly. Let me know how many pages you can handle per day and when you can start; I’m ready to share the first batch right away.