Funda Property Listing Scrape
Budget: €30 – €250 EUR
I need a complete, one-time extraction of all current sales and rental listings found at
https://www.funda.nl/zoeken/koop?availability=[%22negotiations%22,%22unavailable%22,%22available%22]
and
https://www.funda.nl/zoeken/huur?selected_area=[%22nl%22]&availability=[%22available%22,%22negotiations%22,%22unavailable%22]
For every property on those result pages I want the following details captured precisely as they appear on Funda:
• full street address
• postal code
• residential area in m²
• municipality
• listed price
The final dataset must be delivered in a single, well-structured Excel file (.xlsx) with clear column headers and no duplicate rows. A short “read-me” tab or text file that explains any data cleaning or assumptions you had to make will also be appreciated.
Because this is a one-off job, efficient turnaround is important to me. Please outline:
1. the approach and tools you will use (e.g., Python, Scrapy, Selenium, BeautifulSoup, Playwright, etc.) while respecting the site’s pagination and anti-bot measures;
2. the estimated time you need from award to delivery;
3. a realistic fixed price for the full scrape, including any post-processing needed to ensure clean, accurate data.
If you can optionally supply the scraping script as part of the hand-off, note that in your proposal—it’s a plus but not mandatory.
I will review submissions mainly on data accuracy, speed of delivery, and clarity of your proposed method.
https://www.funda.nl/zoeken/koop?availability=[%22negotiations%22,%22unavailable%22,%22available%22]
and
https://www.funda.nl/zoeken/huur?selected_area=[%22nl%22]&availability=[%22available%22,%22negotiations%22,%22unavailable%22]
For every property on those result pages I want the following details captured precisely as they appear on Funda:
• full street address
• postal code
• residential area in m²
• municipality
• listed price
The final dataset must be delivered in a single, well-structured Excel file (.xlsx) with clear column headers and no duplicate rows. A short “read-me” tab or text file that explains any data cleaning or assumptions you had to make will also be appreciated.
Because this is a one-off job, efficient turnaround is important to me. Please outline:
1. the approach and tools you will use (e.g., Python, Scrapy, Selenium, BeautifulSoup, Playwright, etc.) while respecting the site’s pagination and anti-bot measures;
2. the estimated time you need from award to delivery;
3. a realistic fixed price for the full scrape, including any post-processing needed to ensure clean, accurate data.
If you can optionally supply the scraping script as part of the hand-off, note that in your proposal—it’s a plus but not mandatory.
I will review submissions mainly on data accuracy, speed of delivery, and clarity of your proposed method.
Related categories:
Python
Excel
Web Scraping
Software Architecture
Data Mining
Scrapy
Data Extraction
BeautifulSoup