Structured data extraction from E-Mail

Job ID: 39264097

Budget: €30 – €250 EUR

I need to extract data from a daily email. I'm uploading an email I saved today here.

Each email contains 10 or more information blocks (highlighted in gray). The following must be extracted individually from each block (like in the following example):

* Source (example: "Rhein-Zeitung from April 1, 2025")
* Link to newspaper logo ("Rhein-Zeitung" - https://www.genios.de/resource/logos-big/rztg.png)
* Pre-title (see below, e.g., "Schmerzmedikament")
* Title ("Neuwieder Ausstellung inspiriert auch andere Städte Schau zu nicht sichtbaren Beeinträchtigungen geht auf die Reise")
* Hit area ("...Innenstadt, auch in den anderen Stadtteilen könnten Neuwiederinnen und Neuwieder, etwa mit Autismus oder ADHS...")
* Link to the original article (https://bib-voebb.genios.de/em/67eb8f360c46ff2fe2d01cdf/list_1.0.1/doc/RZTG__5a87174d196ec140927b9c56b224c71b0421aecd/1 - secured by paywall)
* Number of words (313)
* Additionally: Date of extraction

and, for example, to be transferred something like an Excel spreadsheet. Note: Not all fields are always filled!

Daily entry is made easier for me with a simple Thunderbird macro or plugin, but a Python script will also work in a pinch.

If our collaboration works, scraping the original could also be discussed.