Large-Scale Web Scraping Project (108K+ Auction Listings) with Ongoing Weekly Work
Budget: $30 – $250 USD
We are looking for an experienced web scraping specialist to extract historical auction data from an ecommerce / auction website.
This project is the first of several large scraping efforts, with the intention of establishing a long-term, ongoing working relationship. Successful completion of this initial project will lead to additional sites to scrape and recurring weekly scraping work as new auctions complete.
Scope of Work
The website contains approximately 108,000 item pages that must be scraped.
All items are indexed and accessible via pagination, with 240 item links per index page.
Index of items can be found here:
https://rb.gy/timd6q
Each indexed link points to an individual item that has already sold in a prior auction.
Data to Be Collected (Per Item)
For each sold item page, you must scrape and return the following fields:
URL
Date Sold
Number of Bids
Sale Price
Item Name
Sport
Front Image URL
Back Image URL
Important:
The image URLs are not directly visible in the HTML and are loaded via JavaScript. You must be able to identify and extract the true image URLs using an advanced crawler or rendering-based approach.
Technical Requirements
Proven experience scraping large datasets (100K+ pages)
Ability to scrape JavaScript-rendered content
Clean, structured, deduplicated output
Respectful scraping practices that avoid unnecessary load
Deliverable
A structured CSV file containing all scraped records
Consistent formatting and clearly labeled columns
Proof of Capability
Before a final freelancer is selected, you will be required to provide a small sample scrape demonstrating:
Accurate field extraction
Correct image URLs
Proper CSV formatting
Ongoing & Future Work
This is not a one-off project.
After successful completion of this scrape:
We have multiple additional auction and ecommerce sites that need to be scraped
We will require ongoing weekly updates as new auctions complete
Our goal is to establish a regular, consistent weekly project with a reliable scraping partner
Preference will be given to freelancers interested in a long-term engagement rather than a single job.
Payment
Payment will be made via Freelancer escrow
Funds released after the full dataset is delivered and verified
How to Apply
Please include:
Your experience with large-scale web scraping
The tools / stack you plan to use
Confirmation you can extract JavaScript-loaded image URLs
Estimated timeline for completion
Confirmation you are open to ongoing weekly work
Only experienced scraping professionals should apply.
This project is the first of several large scraping efforts, with the intention of establishing a long-term, ongoing working relationship. Successful completion of this initial project will lead to additional sites to scrape and recurring weekly scraping work as new auctions complete.
Scope of Work
The website contains approximately 108,000 item pages that must be scraped.
All items are indexed and accessible via pagination, with 240 item links per index page.
Index of items can be found here:
https://rb.gy/timd6q
Each indexed link points to an individual item that has already sold in a prior auction.
Data to Be Collected (Per Item)
For each sold item page, you must scrape and return the following fields:
URL
Date Sold
Number of Bids
Sale Price
Item Name
Sport
Front Image URL
Back Image URL
Important:
The image URLs are not directly visible in the HTML and are loaded via JavaScript. You must be able to identify and extract the true image URLs using an advanced crawler or rendering-based approach.
Technical Requirements
Proven experience scraping large datasets (100K+ pages)
Ability to scrape JavaScript-rendered content
Clean, structured, deduplicated output
Respectful scraping practices that avoid unnecessary load
Deliverable
A structured CSV file containing all scraped records
Consistent formatting and clearly labeled columns
Proof of Capability
Before a final freelancer is selected, you will be required to provide a small sample scrape demonstrating:
Accurate field extraction
Correct image URLs
Proper CSV formatting
Ongoing & Future Work
This is not a one-off project.
After successful completion of this scrape:
We have multiple additional auction and ecommerce sites that need to be scraped
We will require ongoing weekly updates as new auctions complete
Our goal is to establish a regular, consistent weekly project with a reliable scraping partner
Preference will be given to freelancers interested in a long-term engagement rather than a single job.
Payment
Payment will be made via Freelancer escrow
Funds released after the full dataset is delivered and verified
How to Apply
Please include:
Your experience with large-scale web scraping
The tools / stack you plan to use
Confirmation you can extract JavaScript-loaded image URLs
Estimated timeline for completion
Confirmation you are open to ongoing weekly work
Only experienced scraping professionals should apply.
Related categories:
Python
Web Scraping
Software Architecture
Data Mining
Scrapy
Data Scraping
Data Extraction
BeautifulSoup
Selenium
Automation