University Data Scraping to Excel
Budget: ₹12,500 – ₹37,500 INR
I have identified a public website that profiles universities and now need a clean, well-structured Excel file built from it. The task is strictly text-based scraping: every record should capture each university’s courses offered, admission requirements, and contact information exactly as it appears online.
Here is how I intend the work to unfold:
• Crawl or scrape the target site at scale (Python, BeautifulSoup, Scrapy, or a similar tool is fine) and harvest the three data points for every university listed.
• Normalise spelling, line breaks and white-space where necessary, yet keep the original wording intact—no summarising.
• Deliver one .xlsx file: each row equals one university, columns labelled Courses Offered, Admission Requirements and Contact Information.
• Spot-check a sample together to be sure nothing was missed before final hand-off.
Accuracy is paramount; incomplete rows will be returned for correction. Once the sheet passes the spot-check, the job is done.
Here is how I intend the work to unfold:
• Crawl or scrape the target site at scale (Python, BeautifulSoup, Scrapy, or a similar tool is fine) and harvest the three data points for every university listed.
• Normalise spelling, line breaks and white-space where necessary, yet keep the original wording intact—no summarising.
• Deliver one .xlsx file: each row equals one university, columns labelled Courses Offered, Admission Requirements and Contact Information.
• Spot-check a sample together to be sure nothing was missed before final hand-off.
Accuracy is paramount; incomplete rows will be returned for correction. Once the sheet passes the spot-check, the job is done.
Related categories:
Python
Data Entry
Web Scraping
Data Mining
Scrapy
Data Scraping
BeautifulSoup
Data Analysis