Need help on Data Mining task with PySpark

Job ID: 36066445

Budget: $10 – $30 USD

Two python scripts, task1.py and task2.py, that adhere and successfully execute the requirements.

The task is implementing the SON algorithm for finding frequent itemsets using MapReduce. I need to use Python 3.6 and Spark RDD. So far I am able to find the frequent itemsets from the first pass of the SON algorithm, but I am having trouble with formatting and creating the rest of the algorithm.

I need help finishing this task using python and Spark as soon as possible.

thanks.
Related categories: Python Machine Learning (ML) Data Mining Hadoop PySpark