Senior Databricks Data Modeller

Job ID: 40425618

Budget: ₹750 – ₹1,250 INR

This is a six-week, fully remote contract for someone who can jump in immediately and own the data-modelling stream of an ongoing Databricks implementation across India-based teams. My environment already sits on Databricks with Delta Lake and Unity Catalog enabled, and I need a seasoned modeller who can translate business requirements into robust Star and Snowflake schemas, then bring them to life with PySpark and advanced SQL.

You will refine our Medallion architecture (Bronze → Silver → Gold), implement both Type 1 and Type 2 SCD strategies, and tune the pipelines for speed through smart partitioning and other optimisation techniques. The datasets involved are large, structured and semi-structured, so hands-on experience handling such volumes in Databricks is essential.

Key deliverables
• Logical and physical data models documented and version-controlled
• PySpark notebooks / SQL scripts that create the Star and Snowflake tables in Delta Lake under Unity Catalog governance
• Proven SCD Type 1 & 2 routines integrated into the Medallion layers
• Performance benchmark report showing throughput gains from optimisation work
• A short hand-off session (recorded) walking the team through design decisions and next steps

If you have 5–10 years of data-engineering experience, can start right away, and have shipped at least one end-to-end Databricks project featuring Delta Lake and Unity Catalog, I’d love to review your profile and discuss how quickly we can get you onboarded.