Build Separate Web Scraping Tools for Multiple Public Sites
Budget: $30 – $250 USD
I need a developer to collect data from multiple public websites and deliver it in a clean, structured format. This is for legitimate data extraction from publicly available pages. I will share the target URLs and exact data fields with shortlisted candidates.
Scope of work
Scrape data from multiple public websites (details shared after shortlisting)
Extract specific fields consistently and handle pagination/filtering where needed
Normalize/clean the data (remove duplicates, consistent formatting)
Export results to CSV/Excel/JSON (format to be confirmed)
Provide a repeatable solution (script or small app) that I can run on demand
Basic documentation: how to run it, how to adjust settings, where outputs go
Quality requirements
Reliable scraping with error handling and retries
Respectful request rate / throttling to avoid overloading sites
Clear logging (success/fail, pages processed)
Ability to adapt if page structure changes
Experience with Python (Scrapy/BeautifulSoup/Selenium/Playwright) or Node.js
Proxy / rotating user-agents experience (only if needed)
Scheduling/automation (cron, Docker, or cloud run)
Deliverables
Working scraper + instructions
Sample output file(s)
Final dataset from agreed sources (initial run)
To apply, please include
Examples of similar scraping work you’ve done
Scope of work
Scrape data from multiple public websites (details shared after shortlisting)
Extract specific fields consistently and handle pagination/filtering where needed
Normalize/clean the data (remove duplicates, consistent formatting)
Export results to CSV/Excel/JSON (format to be confirmed)
Provide a repeatable solution (script or small app) that I can run on demand
Basic documentation: how to run it, how to adjust settings, where outputs go
Quality requirements
Reliable scraping with error handling and retries
Respectful request rate / throttling to avoid overloading sites
Clear logging (success/fail, pages processed)
Ability to adapt if page structure changes
Experience with Python (Scrapy/BeautifulSoup/Selenium/Playwright) or Node.js
Proxy / rotating user-agents experience (only if needed)
Scheduling/automation (cron, Docker, or cloud run)
Deliverables
Working scraper + instructions
Sample output file(s)
Final dataset from agreed sources (initial run)
To apply, please include
Examples of similar scraping work you’ve done
Related categories:
JavaScript
Python
Data Processing
Web Scraping
Scrapy
Data Extraction
BeautifulSoup
Selenium
Automation