Build a serverless web scraper
Budget: $30 – $250 USD
We are looking for a Python Developer and Web Scraping specialist with knowledge of AWS cloud technologies, especially AWS Lambda. The task is to code a serverless web scraper for an specific web page.
The web scraper will run on its own AWS Lambda Function and connect to other services for queueing, monitoring, and storage. We already have the serverless architecture running, we only need the code for the scraper, running on a Lambda Function, customized for the webpage caracteristics.
The scraper will run at least once a month until completing a scraping task. Every scraping task might contain from 100 to above 100K documents to download, so realiability and mantainability of the scraping process are key factors for this project.
In the attachment you can find more information about the platform and the webpage to scrap
The web scraper will run on its own AWS Lambda Function and connect to other services for queueing, monitoring, and storage. We already have the serverless architecture running, we only need the code for the scraper, running on a Lambda Function, customized for the webpage caracteristics.
The scraper will run at least once a month until completing a scraping task. Every scraping task might contain from 100 to above 100K documents to download, so realiability and mantainability of the scraping process are key factors for this project.
In the attachment you can find more information about the platform and the webpage to scrap