Arabic PDF OCR to Word

Job ID: 40098870

Budget: $10 – $30 USD

I have a batch of Arabic-language PDF documents that need to become clean, fully editable Word files. What matters most to me is text accuracy; every diacritic, word break, and sentence must read exactly as the original. The source scans are of medium quality—some pages look slightly blurry or skewed—so the OCR process will require careful checking rather than a simple one-click export.

Most pages follow standard paragraph formatting, and I’d like that preserved where it doesn’t compromise accuracy. There are only a handful of special symbols that appear throughout the files; please keep them intact. Images are minimal, so you can ignore complex graphic extraction and focus all efforts on perfecting the Arabic text.

Deliverables
• A separate .docx file for each PDF supplied
• Text proof-read and corrected after OCR
• Original page order, headings, and basic layout retained

I’m happy to provide a short sample first so we can verify the recognition quality before you proceed with the full set. Tools such as ABBYY FineReader, Tesseract, Adobe Acrobat, or any other high-calibre OCR workflow are welcome as long as they achieve high fidelity results.

Let me know your estimated turnaround time and the approach you’ll use to guarantee accuracy, and we can get started right away.