PDF scraping research -- 2
Budget: $30 – $250 USD
I am looking for a freelancer who can assist me with PDF scraping research. The ideal candidate should have experience in extracting tables from PDF files. I'm looking for advise not work.
My project is a large scale extraction of tables from PDFs. (thousands of files and a few million pages). I'm looking for advise how I can do it more cheaply and efficiently.
Specific information to extract:
- Automated way to extract tables from the PDFs similar to AWS Textract.
- Is there a way in Python to auto detect tables and extract them like AWS Textract?
I will provide with pdf samples.
Skills and experience required:
- Strong experience in PDF scraping and data extraction
- Proficiency in working with tables and charts
- Knowledge of CSV file format and data manipulation
- Attention to detail and accuracy in extracting and organizing data
My project is a large scale extraction of tables from PDFs. (thousands of files and a few million pages). I'm looking for advise how I can do it more cheaply and efficiently.
Specific information to extract:
- Automated way to extract tables from the PDFs similar to AWS Textract.
- Is there a way in Python to auto detect tables and extract them like AWS Textract?
I will provide with pdf samples.
Skills and experience required:
- Strong experience in PDF scraping and data extraction
- Proficiency in working with tables and charts
- Knowledge of CSV file format and data manipulation
- Attention to detail and accuracy in extracting and organizing data