PDF Extraction Specialist Needed

Job ID: 39302984

Budget: $8 – $15 USD

I need a pdf extraction expert.

I have a hundred documents from which I need to extract text. They all have different and varying formats.

I want to extract it and ensure I don't have orphaned text. What I mean is that on page 5, I might have the page start with text, but the header or subheader for that section might be on page 3, and on page 4, there might be a table or exhibit.

I want to use Python and am open to using gpt-4o to assist
Related categories: Python OpenAI