I would like to web crawl Glassdoor website.
Budget: $30 – $250 SGD
I would like you to scrape all the reviews, ratings, and other relevant information from Glassdoor for around 5,800 U.S. companies that are included in my list (attached), from the very beginning of the Glassdoor website to the very recent period
You need to match the company names in my list with the ID and name in Glassdoor with some name-matching techniques.
The final product should be the csv file that includes company name (in my list), company name, (Glassdoor), company ID (Glassdoor), company gvkey (in my list), review date, reviewer ID, gender, review title, rating, recomend, ceo approval, business outlook, pros, cons, https source, number not helpful, work life balance, culture values, diversity inclusion, career opportunities, compensation benefits, senior management and other relevant information.
After that, you need to share the Python file for web crawling with me so that I can modify some of the parts to customize some other usages.
Thanks and please let me know if you have any questions.
You need to match the company names in my list with the ID and name in Glassdoor with some name-matching techniques.
The final product should be the csv file that includes company name (in my list), company name, (Glassdoor), company ID (Glassdoor), company gvkey (in my list), review date, reviewer ID, gender, review title, rating, recomend, ceo approval, business outlook, pros, cons, https source, number not helpful, work life balance, culture values, diversity inclusion, career opportunities, compensation benefits, senior management and other relevant information.
After that, you need to share the Python file for web crawling with me so that I can modify some of the parts to customize some other usages.
Thanks and please let me know if you have any questions.