Build Separate Web Scraping Tools for Multiple Public Sites

Job ID: 40198130

Budget: $30 – $250 USD

I need a developer to collect data from multiple public websites and deliver it in a clean, structured format. This is for legitimate data extraction from publicly available pages. I will share the target URLs and exact data fields with shortlisted candidates.

Scope of work

Scrape data from multiple public websites (details shared after shortlisting)

Extract specific fields consistently and handle pagination/filtering where needed

Normalize/clean the data (remove duplicates, consistent formatting)

Export results to CSV/Excel/JSON (format to be confirmed)

Provide a repeatable solution (script or small app) that I can run on demand

Basic documentation: how to run it, how to adjust settings, where outputs go

Quality requirements

Reliable scraping with error handling and retries

Respectful request rate / throttling to avoid overloading sites

Clear logging (success/fail, pages processed)

Ability to adapt if page structure changes

Experience with Python (Scrapy/BeautifulSoup/Selenium/Playwright) or Node.js

Proxy / rotating user-agents experience (only if needed)

Scheduling/automation (cron, Docker, or cloud run)

Deliverables

Working scraper + instructions

Sample output file(s)

Final dataset from agreed sources (initial run)

To apply, please include

Examples of similar scraping work you’ve done