Indiegogo Texts Scraping (3839 campaigns)

Job ID: 39862597

Budget: €6 – €12 EUR

I have a CSV file with Indiegogo campaign URLs (attached). I want to extract all publicly available textual content from each campaign page and return it in the same file with new columns for each text type.

Input: attached CSV with campaign URLs in column D.

For each URL, scrape all textual content visible on the campaign page (no images or videos).

Create additional columns for:
- project_description (main story or overview)
- perks_text (text content describing perks, if present)
- updates
- comments
- FAQs
- notes (optional, for failed or duplicate URLs)

Output: same CSV returned with extra columns and UTF-8 encoding. Expected minimum success rate: 95% non-empty text fields.

Requirement: Must have previous experience scraping Indiegogo (please mention the year and sample size of your previous dataset).

Preferred qualifications:
- Demonstrated Indiegogo scraping project or an existing dataset of campaign texts.
- Familiarity with academic or research-quality data collection (clear documentation, reproducibility).

Deliverables:
- One CSV with the new text columns added.
- A short log file listing any failed URLs or pages with empty content.
- (Optional) If you already have a pre-existing dataset of Indiegogo campaign texts from 2018–2025 (3,000+ campaigns, representative sample), you may submit a short description and sample instead of scraping from my URLs.

Timeline: Pilot of 50–100 campaigns first (to validate output), followed by the remaining URLs once approved.

Budget: Flexible for quality work.

Notes: This dataset will be used for an academic research project studying how crowdfunding campaign language evolved over time. I’m only interested in publicly visible, text-based content.
Related categories: Web Scraping Data Mining