Pyspark Developer Needed for Cosine Similarity

Job ID: 38359206

Budget: ₹1,500 – ₹12,500 INR

I'm looking for a talented Pyspark Developer who has experience in working with large datasets and is well-versed in PySpark above version 3.0. The primary task involves creating user-defined function code in PySpark for applying cosine similarity on two text columns.

Key Requirements:
- Handling large datasets (more than 1GB) efficiently
- Proficient in PySpark (above version 3.0)
- Experienced in implementing cosine similarity
- Background in health care data is a plus

Your primary responsibilities will include:
- Writing efficient and scalable code
- Applying cosine similarity on two text columns
- Ensuring the code can handle large datasets

This project is a great opportunity for a Pyspark Developer to showcase their skills in handling big data and implementing complex algorithms.
Related categories: Data Science Microsoft Azure PySpark NLP