Web Scraper for Art Catalog

Job ID: 40592216

Budget: $10 – $30 USD

I have (what would seem to be) a small website scraping project that involves subpages.

The website houses a library of art items. A user can pull up a search page that has all the art for a specific artis. Those search results display a thumbnail and name of the piece which links to a dedicated page for the piece. The dedicated page then displays various data points that need to be extracted. Some pieces also have variations that are accessible through a drop down on the dedicated page. The contents of that drop down vary by piece. So there are 2 layers below the search thumbnail that need to be scraped.

While some data fields are consistent, the data available and displayed for an individual piece may vary. For example, some may identify the manufacturer, event it was released at, or the original price. Others may not have that same information. There is however information that can be identified as consistently inapplicable. Images are not needed (I'm concerned about size).

While I have one artist in mind right now (with about 950-1000 individual pieces), ideally I would like a scraper built that I could reuse to gather the catalog for other artists or that could be used, on-demand style, to gather information for newly added pieces.

I am not a developer or technology professional, so it would need to be available to me through one of the common extensions like Thunderbit, Browse AI, or Data Miner or very user friendly. I am moderately competent in updating inputs like starting URLs, artist name, or dates if logically developed. This is intended for personal use (and maybe a few friends) and will be ultimately cleaned up and maintained in Excel so output as .xls or .csv is a requirement.