Extracting structured data from PDF

Job ID: 31206799

Budget: £250 – £750 GBP

We need to convert PDF documents into structured data format. The PDF is a Confirmation Statement from UK’s companies house and we need names of shareholders and number of shares they hold extracted.

The script should be built in such way so it’s easy to deploy as a serverless app (ie Lambda).
Related categories: Python Data Processing PDF Machine Learning (ML) OCR