PDF Tables to Excel
Budget: $10 – $30 USD
I have a collection of PDFs with tables that need to end up in a clean, well-structured Excel workbook. The twist is that the table layout is not the same from file to file, so a one-size-fits-all script will not work—you’ll need logic that can adapt to shifting column counts, merged cells, and the occasional header tweak.
Here is what I need:
• Accurate extraction of every table, regardless of layout changes.
• A single Excel (.xlsx) file with each source PDF represented on a separate sheet, preserving numeric formats and basic styling where possible.
• A quick note on any tables that cannot be captured automatically so I can review them manually.
I’m open to whatever stack you prefer—Python (pandas, tabula-py, camelot), Power Query, or even a smart VBA solution—as long as the final workbook is reliable and reproducible.
When you respond, please include a sample or brief clip from past work that shows you have already tackled inconsistent PDF tables. That reference will help me move quickly to award the project.
Here is what I need:
• Accurate extraction of every table, regardless of layout changes.
• A single Excel (.xlsx) file with each source PDF represented on a separate sheet, preserving numeric formats and basic styling where possible.
• A quick note on any tables that cannot be captured automatically so I can review them manually.
I’m open to whatever stack you prefer—Python (pandas, tabula-py, camelot), Power Query, or even a smart VBA solution—as long as the final workbook is reliable and reproducible.
When you respond, please include a sample or brief clip from past work that shows you have already tackled inconsistent PDF tables. That reference will help me move quickly to award the project.
Related categories:
JavaScript
Python
Visual Basic
Excel
Scripting
Data Extraction
Data Analysis
Automation