AI/ML PDF Data Extraction Specialist Needed
Budget: $250 – $750 CAD
I'm seeking a skilled professional to extract data from low-quality eFax PDF documents using various tools and techniques. This project primarily involves AI/ML data extraction, focusing mostly on text data.
Key Responsibilities:
- Utilize EasyOCR, pdfplumber, and OpenCV for document processing.
- Implement Next.js with Python, pdf.js, and Label Studio for annotation and user interface.
- Apply spaCy for Natural Language Processing and Named Entity Recognition (NER).
- Leverage PyTorch and Hugging Face's Datasets and Transformers within a machine learning framework.
- Store metadata in PostgreSQL.
Ideal Skills and Experience:
- Proficiency in OCR, PDF processing, and image handling.
- Strong background in Python and Next.js.
- Experience with spaCy, PyTorch, and Hugging Face's Transformers.
- Familiarity with PostgreSQL.
- Ability to work with low-quality documents.
The goal of this project is to facilitate AI/ML data extraction from these PDF documents. I look forward to receiving your bids.
Key Responsibilities:
- Utilize EasyOCR, pdfplumber, and OpenCV for document processing.
- Implement Next.js with Python, pdf.js, and Label Studio for annotation and user interface.
- Apply spaCy for Natural Language Processing and Named Entity Recognition (NER).
- Leverage PyTorch and Hugging Face's Datasets and Transformers within a machine learning framework.
- Store metadata in PostgreSQL.
Ideal Skills and Experience:
- Proficiency in OCR, PDF processing, and image handling.
- Strong background in Python and Next.js.
- Experience with spaCy, PyTorch, and Hugging Face's Transformers.
- Familiarity with PostgreSQL.
- Ability to work with low-quality documents.
The goal of this project is to facilitate AI/ML data extraction from these PDF documents. I look forward to receiving your bids.