German Gastronomy Contact Data Scrape
Budget: $250 – $750 USD
I need a fresh, well-structured dataset of German gastronomy venues—specifically restaurants, cafés and bars—complete with the name of the restaurant, the owner’s name, a working e-mail address, postal code and town. The fastest way to achieve scale here is web scraping, so please rely on your favourite Python stack (Scrapy, BeautifulSoup, Selenium or a comparable tool) rather than manual look-ups.
Scope
• Nationwide coverage: all federal states in Germany are relevant.
• One line per venue, no duplicates.
• Contact person must be the owner or manager; other roles are not required.
• Data sources must be publicly available pages (official websites, online menus, Imprint pages, Google Maps, etc.) to stay GDPR-compliant.
Deliverables
1. CSV or Excel file containing: Business Name, Owner Name, E-mail, Postal Code, Town, Source URL.
2. A short README explaining the scraping approach and any limitations.
3. Verification sample: 50 random entries re-checked for accuracy ≥ 90 %.
I will review the sample first; once approved you can scale to the full list. Please clarify the volume you expect to deliver and your estimated timeline when you respond.
We expect about 50.000 - 100.000 datasets.
Scope
• Nationwide coverage: all federal states in Germany are relevant.
• One line per venue, no duplicates.
• Contact person must be the owner or manager; other roles are not required.
• Data sources must be publicly available pages (official websites, online menus, Imprint pages, Google Maps, etc.) to stay GDPR-compliant.
Deliverables
1. CSV or Excel file containing: Business Name, Owner Name, E-mail, Postal Code, Town, Source URL.
2. A short README explaining the scraping approach and any limitations.
3. Verification sample: 50 random entries re-checked for accuracy ≥ 90 %.
I will review the sample first; once approved you can scale to the full list. Please clarify the volume you expect to deliver and your estimated timeline when you respond.
We expect about 50.000 - 100.000 datasets.
Related categories:
Python
Web Scraping
Software Architecture
Data Mining
Scrapy
BeautifulSoup
Selenium
Data Collection