One-Time Website Text Scrape

Job ID: 39908822

Budget: $10 – $30 USD

I need the visible text from a single website captured and delivered in an Excel file. This is a one-time job—no scheduling or ongoing automation required.

Scope
• Access the URL I provide and extract all relevant on-page text (headings, paragraphs, table cells, and lists).
• Ignore images, PDFs, or embedded media; text only.
• Consolidate the cleaned content into a well-structured worksheet with clear column headers (e.g., Page URL, Section, Extracted Text).

Technical Notes
A lightweight approach using Python, BeautifulSoup, Scrapy, or a similar tool is fine as long as the final XLSX is clean, deduplicated, and readable. No API work is expected.

Acceptance
I’ll consider the task complete once I can open the Excel file, see every page/section represented without missing text, and cross-check a random sample against the live site.

If you can start right away and finish quickly, let me know your turnaround time and any clarifications you need.