Convert Scanned pdf book to doc using Google vision OCR

Job ID: 32202774

Budget: $250 – $750 USD

Looking for someone to build Python software will take old scanned pdf book with mix of multi languages text and images and run it through google vision OCR to digitalize all the text.
* Software must be able to perform OCR on a whole book or just few pages for testing
*Software need to be a multi-threaded , have the option to bulk OCR all files in a folder.
*Software must be able to handle possible interruptions and able to continue the work from where it stopped.
*Software must be able to keep the structure of the book like paragraphs and images locations for example.
*Software must be able to reduce the size of the book since it's digitizing all the text and not just adding a text layer.
* And Finally would prefer if there is a user GUI.
Related categories: PHP JavaScript Python Software Architecture OCR