Web scraping from html page
Budget: ₹1,500 – ₹12,500 INR
We are seeking a skilled web scraping expert to scrape html pages and convert them into a pdf file. The selected candidate will be responsible for creating a web scraper script that can automatically extract and organize data from various websites. The data should be neatly formatted into a pdf document for easy accessibility.
The website is https://incometaxindia.gov.in/pages/acts/income-tax-act.aspx in which there are various links - section - 1, section - 2, etc. The text which is present when a link is opened is to be scraped. There may be an odd number as well, like section - 5A. Also, there are pages which will give more links/sections (there are upto 93 pages).
When the section is opened, there may be footnotes. The text when the footnote is opened is to be scraped if possible.
The website is https://incometaxindia.gov.in/pages/acts/income-tax-act.aspx in which there are various links - section - 1, section - 2, etc. The text which is present when a link is opened is to be scraped. There may be an odd number as well, like section - 5A. Also, there are pages which will give more links/sections (there are upto 93 pages).
When the section is opened, there may be footnotes. The text when the footnote is opened is to be scraped if possible.