Scrape Website Content and Images

Job ID: 35863624

Budget: $12 – $50 SGD

I wanted to scrape these links from this website (Dot Property)

Links:
https://www.dotproperty.com.ph/houses-for-sale/metro-manila/quezon-city
https://www.dotproperty.com.ph/houses-for-sale/cebu/cebu-city
https://www.dotproperty.com.ph/houses-for-sale/metro-manila/makati
https://www.dotproperty.com.ph/houses-for-sale/metro-manila/caloocan
https://www.dotproperty.com.ph/houses-for-sale/metro-manila/manila
https://www.dotproperty.com.ph/houses-for-sale/metro-manila/mandaluyong
https://www.dotproperty.com.ph/houses-for-sale/metro-manila/marikina
https://www.dotproperty.com.ph/houses-for-sale/metro-manila/muntinlupa
https://www.dotproperty.com.ph/houses-for-sale/metro-manila/pasay
https://www.dotproperty.com.ph/houses-for-sale/metro-manila/pasig
https://www.dotproperty.com.ph/houses-for-sale/metro-manila/san-juan
https://www.dotproperty.com.ph/houses-for-sale/metro-manila/taguig
https://www.dotproperty.com.ph/houses-for-sale/metro-manila/valenzuela
https://www.dotproperty.com.ph/houses-for-sale/laguna/santa-rosa
https://www.dotproperty.com.ph/houses-for-sale/laguna/calamba
https://www.dotproperty.com.ph/houses-for-sale/laguna/san-pablo
https://www.dotproperty.com.ph/houses-for-sale/laguna/san-pedro
https://www.dotproperty.com.ph/houses-for-sale/batangas/batangas-city
https://www.dotproperty.com.ph/houses-for-sale/batangas/lipa
https://www.dotproperty.com.ph/houses-for-sale/pampanga/angeles
https://www.dotproperty.com.ph/houses-for-sale/pampanga/porac
https://www.dotproperty.com.ph/houses-for-sale/pampanga/san-fernando
https://www.dotproperty.com.ph/houses-for-sale/pampanga/mabalacat

The the contents that needs to be scrape are:

1. Ad Title
2. Description
3. Price
4. Images (filename in comma delimited) (see attached csv file as example)
- Original images needs to be saved locally.
5. Seller Name
6. Mobile Number

I have highlighted in green (in House-And-Lot-For-Sale-Makati-City.csv) the data that is needed, the not highlighted can leave it blank.

Also links above (ad-listing.jpg is an example) has all the ads that needs to be scraped but needs to go inside to the ad (see ad-page.jpg) to get the other contents.

Also to take note that the images download needs to be renamed in a random filename so it won't get conflict or overwritten. (see the .csv file under images to get what I mean).


You can scrape and provide me the data... or probably a script to run this on a server. Lets chat/discuss.
Related categories: Python Web Scraping Scrapy Data Scraping