Scrape 285 URLs to CSV
Budget: $30 – $250 AUD
I have 285 public-facing project pages and need the same 26 data points lifted from each one. You will receive a spreadsheet with every URL and a clear field dictionary so you can move straight to extraction with Python, BeautifulSoup, Scrapy, or whichever tool chain you prefer.
The finished dataset must come back as a single Excel/CSV file. Before you hand it over, give it a quick polish: apply basic, uniform formatting, drop any duplicates, and make sure each column lines up with the field names I supply. No heavy ETL work—just that first-pass cleanup so I can analyse the file immediately.
Deliverables
• CSV (or XLSX) containing 13 columns × ~285 rows, fully populated where data exists
• Basic formatting and de-duplication applied
• Short note flagging any URLs or fields that could not be captured
A quick turnaround is ideal; the job should be straightforward for anyone comfortable with web scraping and light data wrangling.
We will provide the file for review to applicants.
The finished dataset must come back as a single Excel/CSV file. Before you hand it over, give it a quick polish: apply basic, uniform formatting, drop any duplicates, and make sure each column lines up with the field names I supply. No heavy ETL work—just that first-pass cleanup so I can analyse the file immediately.
Deliverables
• CSV (or XLSX) containing 13 columns × ~285 rows, fully populated where data exists
• Basic formatting and de-duplication applied
• Short note flagging any URLs or fields that could not be captured
A quick turnaround is ideal; the job should be straightforward for anyone comfortable with web scraping and light data wrangling.
We will provide the file for review to applicants.
Related categories:
Python
Data Processing
Data Entry
Excel
Web Scraping
Data Mining
Scrapy
Data Extraction
BeautifulSoup
Data Collection