PDF scraping research -- 2

Job ID: 36928799

Budget: $30 – $250 USD

I am looking for a freelancer who can assist me with PDF scraping research. The ideal candidate should have experience in extracting tables from PDF files. I'm looking for advise not work.

My project is a large scale extraction of tables from PDFs. (thousands of files and a few million pages). I'm looking for advise how I can do it more cheaply and efficiently.

Specific information to extract:

- Automated way to extract tables from the PDFs similar to AWS Textract.
- Is there a way in Python to auto detect tables and extract them like AWS Textract?

I will provide with pdf samples.

Skills and experience required:

- Strong experience in PDF scraping and data extraction
- Proficiency in working with tables and charts
- Knowledge of CSV file format and data manipulation
- Attention to detail and accuracy in extracting and organizing data
Related categories: Python Data Processing PDF Data Mining Big Data Sales