Web Scraping, Data Extraction & Excel Automation Expert Needed (Chemical Industry)
Budget: ₹1,500 – ₹12,500 INR
Hi,
We are a chemical trading company based in Gujarat, India. We are looking for a skilled freelancer to help us automate data collection and build a structured database of chemical products from multiple sources.
Currently, we collect and manage data manually from websites, PDFs, and Excel files. We want to automate this entire process and create a centralized system.
Project Requirements:
1. Web Scraping
Scrape data from multiple chemical supplier websites
Extract:
Product Name
CAS Number
Price
Pack Size
Specifications (if available)
2. Data from Existing Files
Extract and process data from:
PDF files
Existing Excel sheets
Convert unstructured data into a clean, usable format
3. Data Consolidation (Very Important)
Merge all data sources:
Website data
PDF data
Excel data
Combine and organize data:
Category-wise
CAS Number-wise (primary key)
Remove duplicates and standardize entries
4. Excel Output (Must Requirement)
Final data should be delivered in a well-structured Excel format
Excel should be:
Clean and easy to use
Filterable and searchable
Structured based on our business needs
5. Database Storage
Store all data in a structured database (MySQL / PostgreSQL / MongoDB preferred)
6. Automation (Important)
System should:
Automatically update website data daily or periodically
Allow easy re-import of new PDF/Excel files
Should be scalable for adding more websites in future
We are a chemical trading company based in Gujarat, India. We are looking for a skilled freelancer to help us automate data collection and build a structured database of chemical products from multiple sources.
Currently, we collect and manage data manually from websites, PDFs, and Excel files. We want to automate this entire process and create a centralized system.
Project Requirements:
1. Web Scraping
Scrape data from multiple chemical supplier websites
Extract:
Product Name
CAS Number
Price
Pack Size
Specifications (if available)
2. Data from Existing Files
Extract and process data from:
PDF files
Existing Excel sheets
Convert unstructured data into a clean, usable format
3. Data Consolidation (Very Important)
Merge all data sources:
Website data
PDF data
Excel data
Combine and organize data:
Category-wise
CAS Number-wise (primary key)
Remove duplicates and standardize entries
4. Excel Output (Must Requirement)
Final data should be delivered in a well-structured Excel format
Excel should be:
Clean and easy to use
Filterable and searchable
Structured based on our business needs
5. Database Storage
Store all data in a structured database (MySQL / PostgreSQL / MongoDB preferred)
6. Automation (Important)
System should:
Automatically update website data daily or periodically
Allow easy re-import of new PDF/Excel files
Should be scalable for adding more websites in future
Related categories:
Data Entry
Excel
Web Scraping
MySQL
Data Mining
Data Extraction
Automation
Database Management