Parallelize COP-K-MEANS algorithm using exclusively PYSPARK

Job ID: 34063267

Budget: €30 – €250 EUR

I want a person that understand 100% the cop-k-means algorithm. As you know, it's an iterative algorithm and it takes a lot of time to calculate a big amount of data. The purpose of this project is to program the algorithm in pyspark to make it able to BIG DATA (using spark functions and paralleling processes). Also, the efficiency and the validation of constraints is important for this problem.