Scanned Documents Data Tabulation

Job ID: 39906747

Budget: ₹600 – ₹1,500 INR

I have a collection of scanned documents saved as high-resolution JPG files. Each image contains both text and numbers that must be captured exactly and placed into a structured spreadsheet. I will provide a column template that shows the order of fields—things like Document ID, Dates, Names, descriptive text fields, and the various numeric values present on the page.

Your job is to run reliable OCR, proof-read the results, and copy everything into the template, applying a few straightforward formatting rules along the way: strip currency symbols while keeping the figures, convert all dates to DD-MM-YYYY, and make sure multi-line remarks are neatly wrapped inside their cell instead of spilling over. Any toolset is fine—Tesseract, Adobe Acrobat, Excel, Google Sheets—so long as the final file opens cleanly in .xlsx and meets the accuracy checks.

Deliverables
• Completed .xlsx file with every piece of text and every number from the documents
• Optional OCR/Text output you generated during processing
• Short note flagging any areas that were unreadable or ambiguous

I will verify the spreadsheet through random spot-checks; payment is released once every field matches the source and the formatting rules are followed.