Arabic PDF Text Extraction

Job ID: 39837939

Budget: $250 – $750 USD

I have several PDFs written entirely in Arabic and I simply need every word pulled out and placed into a single Microsoft Word document. The files contain only text—no images or tables—so you don’t have to recreate any complex layout. A clean, continuous flow of Arabic text is all that matters; replicating the original fonts, columns, or styling isn’t required.

Accuracy is critical. I’ll quickly spot missing diacritics, broken ligatures, or right-to-left issues, so please rely on dependable OCR or direct extraction tools such as Adobe Acrobat, ABBYY FineReader, or a well-trained Tesseract setup, then proofread before handing the file over.

Deliverable
• One .docx file per source PDF containing the complete Arabic text in reading order.

I’ll share a sample PDF once we start. If you can return a short test page that looks perfect, the remaining files will follow immediately.