25k Book Metadata Extraction
Budget: $30 – $250 USD
I need a structured database of 25,000 unique titles, pulling data from both Amazon and Goodreads and merging the results into a single, clean file. For every book, capture the following fields exactly:
• Title and subtitle
• ASIN and ISBN
• Number of pages
• Average rating
• Number of reviews
• Publication date
• Full description
• Genre information
• Amazon product URL
• Cover image URL
I’m working on a tight timeline—delivery is needed as soon as possible—so efficient scraping or API techniques (Python, BeautifulSoup, Selenium, Amazon Product Advertising API, Goodreads API, etc.) will be essential. The final hand-off should be a CSV or SQL dump that I can import directly, with no duplicates and consistent formatting across all fields.
Please quote a clear rate per book and give a realistic turnaround estimate when you reply.
• Title and subtitle
• ASIN and ISBN
• Number of pages
• Average rating
• Number of reviews
• Publication date
• Full description
• Genre information
• Amazon product URL
• Cover image URL
I’m working on a tight timeline—delivery is needed as soon as possible—so efficient scraping or API techniques (Python, BeautifulSoup, Selenium, Amazon Product Advertising API, Goodreads API, etc.) will be essential. The final hand-off should be a CSV or SQL dump that I can import directly, with no duplicates and consistent formatting across all fields.
Please quote a clear rate per book and give a realistic turnaround estimate when you reply.
Related categories:
Python
Web Scraping
Software Architecture
Data Mining
Data Scraping
BeautifulSoup
Selenium
API Development