Custom Crawler for Blog Articles

Job ID: 38286981

Budget: $30 – $250 USD

I'm in need of a custom web crawler that can scrape the content from blog websites and help centers using a regex-based system.

Key Requirements:
- The crawler should be able to extract the title, author, publication date, and the content of the articles.
- The content extraction should be specifically customized for blog websites and help centers.

Additional Information:
- The project is of high priority and I'm looking for a quick completion, ideally ASAP.
- Please only apply if you're experienced in building web crawlers and have a strong understanding of regex.

Skills required:
- Proficiency in web scraping and building custom web crawlers.
- Strong knowledge of regex.
- Ability to deliver in a short period.


build it with python / typescript
Related categories: JavaScript Python Web Scraping Data Mining