Build an application for web scraper
Budget: ₹3,000 – ₹4,000 INR
Note: the photo which I have uploaded it’s an chrome extension but I need an application.
I need an application that can instantly pull contact details from university websites in any country and drop them straight into an Excel workbook. The scraper must capture the following for each person it finds: email address, phone number, social media profile links, full name, affiliation, department, and the exact profile URL the data came from.
My ideal workflow is simple: I point the tool at a single domain or a list of university URLs, press start, and receive a neatly formatted .xlsx file. Because institutional sites vary widely—some rely on pagination, others on JavaScript-rendered staff directories—the program should tackle both static and dynamic pages, handle moderate rate-limits, and avoid duplicating records.
I’m flexible on the tech stack; Python with BeautifulSoup/Scrapy, Node with Puppeteer, or another robust approach is fine as long as the finished package runs on a standard Windows machine and is easy to maintain.
Deliverables
• Stand-alone executable (plus readable source code)
• Sample Excel file showing all required fields for at least one university
• Brief README that covers setup, usage, and how to add new domains or refine selectors
If you can build a tool that meets these points and produces reliable, up-to-date university contact lists on demand, let’s get started.
I need an application that can instantly pull contact details from university websites in any country and drop them straight into an Excel workbook. The scraper must capture the following for each person it finds: email address, phone number, social media profile links, full name, affiliation, department, and the exact profile URL the data came from.
My ideal workflow is simple: I point the tool at a single domain or a list of university URLs, press start, and receive a neatly formatted .xlsx file. Because institutional sites vary widely—some rely on pagination, others on JavaScript-rendered staff directories—the program should tackle both static and dynamic pages, handle moderate rate-limits, and avoid duplicating records.
I’m flexible on the tech stack; Python with BeautifulSoup/Scrapy, Node with Puppeteer, or another robust approach is fine as long as the finished package runs on a standard Windows machine and is easy to maintain.
Deliverables
• Stand-alone executable (plus readable source code)
• Sample Excel file showing all required fields for at least one university
• Brief README that covers setup, usage, and how to add new domains or refine selectors
If you can build a tool that meets these points and produces reliable, up-to-date university contact lists on demand, let’s get started.
Related categories:
JavaScript
Python
Excel
Web Scraping
Data Mining
Node.js
App Developer
Scrapy
BeautifulSoup