Credit Card PDFs to Excel
Budget: $250 – $750 USD
I have approximately 120 credit-card statement PDFs that I need converted into a clean, usable Excel sheet. The only information I want pulled is the text—no images or embedded graphics—so you can ignore logos, headers, and other non-text elements.
The statements follow a consistent layout typical of monthly credit-card summaries. Basic accuracy is all that’s required: if the OCR occasionally misses a comma or a line break, that’s fine as long as every transaction line and key summary figure appears in the spreadsheet and sits in its own sensible column. No secondary, line-by-line verification is expected.
Deliverables
• One Excel file per PDF (or a single consolidated workbook, if that’s easier for you) containing all extracted text.
Feel free to use any reliable toolset—Adobe Acrobat, ABBYY FineReader, Python + Tesseract, or your preferred OCR workflow—so long as the final spreadsheet opens cleanly in Microsoft Excel.
The statements follow a consistent layout typical of monthly credit-card summaries. Basic accuracy is all that’s required: if the OCR occasionally misses a comma or a line break, that’s fine as long as every transaction line and key summary figure appears in the spreadsheet and sits in its own sensible column. No secondary, line-by-line verification is expected.
Deliverables
• One Excel file per PDF (or a single consolidated workbook, if that’s easier for you) containing all extracted text.
Feel free to use any reliable toolset—Adobe Acrobat, ABBYY FineReader, Python + Tesseract, or your preferred OCR workflow—so long as the final spreadsheet opens cleanly in Microsoft Excel.
Related categories:
Python
Data Entry
Excel
PDF
LaTeX
OCR
Data Extraction
Data Management
Text Recognition