Looking for an OCR specialist (optical character recognition)
Budget: $30 – $250 USD
I am looking for a person who has vast experience using OCR software (optical character recognition software) and who can help me choose a solution for my needs.
I have many PDF documents that need to be recognized (OCR) and converted into delimited text/ASCII format so that they can be loaded into a database on a Windows machine, in a Windows Server environment. The PDF documents are lists of data. The PDF files all have essentially the same format, with very small variations.
The final solution will be called from a management system (spawning a DOS box) and run either via an executable or a batch file, for example:
C:\> ocr.exe input.pdf output.txt
or
C:\> ocr.bat input.pdf output.txt
The input.pdf file is the PDF that needs to be scanned and converted into text, then sent to the output.txt file.
Here is what I want you to do:
1) I will supply several test PDF documents;
2) You will choose the correct OCR tool to do the job;
3) You will create a program or script to execute the task (via .exe or .bat);
4) You will provide the program or script to me to test;
5) If it works, you will document how it works;
6) Then you get paid;
I understand that most OCR solutions require you to tweek or configure them to be able to better understand the content of the PDF file. You will need to explain this in your documentation.
Please read and understand the project details carefully. Your bid on this project is your final bid. If you are awarded the project you cannot ask for more money or a tip after the project is awarded. You will be paid what you bid. If you have any questions, please ask them before you bid. Your level of professionalism will determine if I do future work with you, as this is the first phase of a multi-phase project.
Thank you
I have many PDF documents that need to be recognized (OCR) and converted into delimited text/ASCII format so that they can be loaded into a database on a Windows machine, in a Windows Server environment. The PDF documents are lists of data. The PDF files all have essentially the same format, with very small variations.
The final solution will be called from a management system (spawning a DOS box) and run either via an executable or a batch file, for example:
C:\> ocr.exe input.pdf output.txt
or
C:\> ocr.bat input.pdf output.txt
The input.pdf file is the PDF that needs to be scanned and converted into text, then sent to the output.txt file.
Here is what I want you to do:
1) I will supply several test PDF documents;
2) You will choose the correct OCR tool to do the job;
3) You will create a program or script to execute the task (via .exe or .bat);
4) You will provide the program or script to me to test;
5) If it works, you will document how it works;
6) Then you get paid;
I understand that most OCR solutions require you to tweek or configure them to be able to better understand the content of the PDF file. You will need to explain this in your documentation.
Please read and understand the project details carefully. Your bid on this project is your final bid. If you are awarded the project you cannot ask for more money or a tip after the project is awarded. You will be paid what you bid. If you have any questions, please ask them before you bid. Your level of professionalism will determine if I do future work with you, as this is the first phase of a multi-phase project.
Thank you