LinkedIn and Glassdoor Data Scraping

Job ID: 38372171

Budget: ₹12,500 – ₹37,500 INR

I'm in need of a seasoned web scraper who can efficiently extract data from LinkedIn and Glassdoor. The target data points are:

- All accessible data related to companies and founders, and other entities of interest given a URL.

This project is majorly open to companies from all industries, so a comprehensive extraction technique would be required. The extraction should be thorough, including start-ups to well-established corporations.

The preferred format for this database should be JSON (JavaScript Object Notation) for easier manipulation and integration with other parts of our system.

Ideal Experience and skills:

- Proven experience in web scraping, especially from LinkedIn and Glassdoor
- Must be using Python as language for scraping.
- Strong understanding of JSON format data management
- Detail-oriented with a knack for capturing the most minute details
- Ability to work within tight deadlines while maintaining high-quality outputs.

If you have the skills and can handle this level of scraping, feel free to place a bid.

Please note: We will own the code and the rights of it. The measure of success will be how robust is the code and able to extract as much details as possible. We will run the Dockerfile and it should be able to setup all anti-scraping measures to prevent blocking.
You must also be able to handle complexities around anti-scraping measures like CAPTCHA, IP blocking, and frequent changes to the website's structure.
Related categories: Python Web Scraping Data Mining Selenium Automation