Academic Email Extraction from PDF Journals

Job ID: 39821578

Budget: ₹600 – ₹1,500 INR

Project Title: Extract 20,000 Corresponding Author Emails from Journal PDFs

Description:
I need a freelancer with experience in PDF data extraction, text mining, and academic data processing to extract 20,000 corresponding author email addresses from journal article PDFs. The PDFs contain multiple volumes/issues of research papers, and the email addresses are usually listed in the author details section.

Requirements:

Extract only corresponding author emails (not all co-authors).

Deliver results in Excel/CSV format with the following columns:

Article Title

Corresponding Author Name

Email Address

Journal Volume/Issue (if available)

Ensure accuracy and no duplicates.

Minimum 20,000 unique corresponding author emails within 5 days.

Skills Needed:

Strong background in data scraping / extraction

Familiarity with academic journal formats (PubMed, Elsevier, Taylor & Francis, Springer, etc.)

Ability to handle large batches of PDFs quickly and accurately

Delivery:

Clean, deduplicated Excel/CSV file with at least 20,000 verified corresponding author emails.

Deadline: 5 days

Budget: Open to negotiation depending on speed, accuracy, and past experience.