PySpark Exercise

Job ID: 31522308

Budget: $10 – $30 USD

I need solutions for the problems in the below attached document in PySpark. I am attaching the solution to first problem and if that is correct, please do the remaining or fix this one as well (in the document) and need for the remaining 3 problems. The solution uses S3 bucket to read the data from, but, you can read it from local env as well. Also, attached the datasets.
Related categories: Python PySpark