Python Web Scraping for Movie Scripts

Job ID: 39168799

Budget: $30 – $250 USD

I'm looking for a Python expert to develop a web scraping script. The script should download movie script PDFs and specific metadata from certain websites.

Requirements:
- Extract movie scripts and metadata from https://www.simplyscripts.com/ and https://www.filmcompanion.in/companion-zone/scripts
- The metadata to be extracted includes: Title, Release date, Genre, IMDB rating.
- Catch and remove duplicates based on the Title.
- Save the extracted metadata in CSV format.
- Please keep the downloaded file in the following format: year_title.pdf
- I should be able to run it locally on my machine. Make sure to handle rate limiting issues if any. If. dependency management is needed then use Poetry(https://python-poetry.org/)

Ideal skills and experience:
- Proficient in Python and web scraping.
- Familiarity with handling and processing PDF files.
- Previous experience in data extraction from movie script websites is a plus.
- Able to format and organize data clearly in CSV files.
Related categories: Python Web Scraping