Optimize Python Web Scraper - Expert
Budget: $10 – $30 USD
I'm seeking an experienced Python developer to review and enhance my existing web scraping code. The current setup scrapes text data from financial institutions' websites and feeds it into a GenAI model. However, I'm encountering restrictions that hinder data extraction.
Requirements:
- Review current code
- Suggest and implement better scraping methods
- Maintain or improve data accuracy and efficiency
- Ensure compliance with website terms of service
Skills and Experience:
- Proficient in Python
- Expertise in web scraping (BeautifulSoup, Scrapy, etc.)
- Familiarity with browser automation (Selenium)
- Experience with handling IP rotation and proxies
- Knowledge of financial websites is a plus
Deliverables:
- Revised, optimized scraping code
- Documentation of changes and improvements
- Recommendations for future maintenance and scalability
More details:
I already have a Python code, which will scrape the data from the websites and feed the data for GenAI to extract the information. I have used multiple options (fallback) to scrape the data, some times due to restrictions from banks' sites, I am not able to extract the content from the websites.
I want to explore an export option to review the code and determine if there is a better way to extract / scrape the data.
Requirements:
- Review current code
- Suggest and implement better scraping methods
- Maintain or improve data accuracy and efficiency
- Ensure compliance with website terms of service
Skills and Experience:
- Proficient in Python
- Expertise in web scraping (BeautifulSoup, Scrapy, etc.)
- Familiarity with browser automation (Selenium)
- Experience with handling IP rotation and proxies
- Knowledge of financial websites is a plus
Deliverables:
- Revised, optimized scraping code
- Documentation of changes and improvements
- Recommendations for future maintenance and scalability
More details:
I already have a Python code, which will scrape the data from the websites and feed the data for GenAI to extract the information. I have used multiple options (fallback) to scrape the data, some times due to restrictions from banks' sites, I am not able to extract the content from the websites.
I want to explore an export option to review the code and determine if there is a better way to extract / scrape the data.
Related categories:
Python
Data Processing
Web Scraping
Django
Software Architecture
Scrapy
Data Extraction
BeautifulSoup
Documentation
Selenium