Scrap web pages in txt format
Budget: $30 – $250 USD
Hi. I would like to scrap all types of websites (built on Wordpress and other softwares) for text-based content (not pdf, video, audio, images, etc) and save them in txt format. I don't want it saved in html format because I do not want files with html codes. I am only interested in the content. The limit of the scope will be stored in a text file (e.g. only scrap content in domains containing the words "aaa", "bbb", etc.) All the txt files are to be scraped and stored in one folder (i.e. don't follow the folder hierarchy of the website).