Expert Web Scraper Required -- 2
Budget: $250 – $750 USD
Project Overview:
We're seeking a highly skilled web scraping expert to design and implement robust scraping solutions for extracting data from a variety of websites. These sites may include dynamic content, require session/cookie management or advanced fingerprinting evasion techniques.
Responsibilities:
Build custom scraping pipelines using tools like Python (Scrapy, Playwright, Selenium), Node.js, or Go.
Handle difficult sites (eg, Cloudflare, PerimeterX, etc.) using rotating proxies, CAPTCHA solvers, headless browsers, or low-level TLS/client fingerprinting.
Design resilient scrapers that can recover from failures, scale efficiently, and adapt to layout/API changes.
Optimize for performance and low detection risk (e.g. rate-limiting, human-like interaction).
Deliver clean, structured output (JSON/CSV/DB) and integrate with APIs or storage systems.
Preferred Skills & Experience:
Deep understanding of HTTP(S), web protocols, and browser behavior.
Experience with
Playwright / Puppeteer
uTLS / JA3 spoofing / TLS fingerprinting
Residential proxy networks
Headless browser stealth techniques
Solid programming skills in Python, Go, or Node.js
Knowledge of data storage systems (PostgreSQL, MongoDB, AWS S3, etc.)
Experience with containerized deployments (Docker, Kubernetes a plus)
Bonus Points:
Experience scraping data from difficult targets
Knowledge of serverless architecture for scalable scraping
Work Setup:
Remote
Flexible hours with weekly progress updates
This can be a one-off project with potential for long-term collaboration if successful
Budget:
Open to hourly or fixed-price proposals based on scope and experience.
We're seeking a highly skilled web scraping expert to design and implement robust scraping solutions for extracting data from a variety of websites. These sites may include dynamic content, require session/cookie management or advanced fingerprinting evasion techniques.
Responsibilities:
Build custom scraping pipelines using tools like Python (Scrapy, Playwright, Selenium), Node.js, or Go.
Handle difficult sites (eg, Cloudflare, PerimeterX, etc.) using rotating proxies, CAPTCHA solvers, headless browsers, or low-level TLS/client fingerprinting.
Design resilient scrapers that can recover from failures, scale efficiently, and adapt to layout/API changes.
Optimize for performance and low detection risk (e.g. rate-limiting, human-like interaction).
Deliver clean, structured output (JSON/CSV/DB) and integrate with APIs or storage systems.
Preferred Skills & Experience:
Deep understanding of HTTP(S), web protocols, and browser behavior.
Experience with
Playwright / Puppeteer
uTLS / JA3 spoofing / TLS fingerprinting
Residential proxy networks
Headless browser stealth techniques
Solid programming skills in Python, Go, or Node.js
Knowledge of data storage systems (PostgreSQL, MongoDB, AWS S3, etc.)
Experience with containerized deployments (Docker, Kubernetes a plus)
Bonus Points:
Experience scraping data from difficult targets
Knowledge of serverless architecture for scalable scraping
Work Setup:
Remote
Flexible hours with weekly progress updates
This can be a one-off project with potential for long-term collaboration if successful
Budget:
Open to hourly or fixed-price proposals based on scope and experience.
Related categories:
Python
Web Scraping
Django
NoSQL Couch & Mongo
Scrapy
Data Extraction
Selenium
API Integration