Senior Databricks Data Modeller
Budget: ₹750 – ₹1,250 INR
This is a six-week, fully remote contract for someone who can jump in immediately and own the data-modelling stream of an ongoing Databricks implementation across India-based teams. My environment already sits on Databricks with Delta Lake and Unity Catalog enabled, and I need a seasoned modeller who can translate business requirements into robust Star and Snowflake schemas, then bring them to life with PySpark and advanced SQL.
You will refine our Medallion architecture (Bronze → Silver → Gold), implement both Type 1 and Type 2 SCD strategies, and tune the pipelines for speed through smart partitioning and other optimisation techniques. The datasets involved are large, structured and semi-structured, so hands-on experience handling such volumes in Databricks is essential.
Key deliverables
• Logical and physical data models documented and version-controlled
• PySpark notebooks / SQL scripts that create the Star and Snowflake tables in Delta Lake under Unity Catalog governance
• Proven SCD Type 1 & 2 routines integrated into the Medallion layers
• Performance benchmark report showing throughput gains from optimisation work
• A short hand-off session (recorded) walking the team through design decisions and next steps
If you have 5–10 years of data-engineering experience, can start right away, and have shipped at least one end-to-end Databricks project featuring Delta Lake and Unity Catalog, I’d love to review your profile and discuss how quickly we can get you onboarded.
You will refine our Medallion architecture (Bronze → Silver → Gold), implement both Type 1 and Type 2 SCD strategies, and tune the pipelines for speed through smart partitioning and other optimisation techniques. The datasets involved are large, structured and semi-structured, so hands-on experience handling such volumes in Databricks is essential.
Key deliverables
• Logical and physical data models documented and version-controlled
• PySpark notebooks / SQL scripts that create the Star and Snowflake tables in Delta Lake under Unity Catalog governance
• Proven SCD Type 1 & 2 routines integrated into the Medallion layers
• Performance benchmark report showing throughput gains from optimisation work
• A short hand-off session (recorded) walking the team through design decisions and next steps
If you have 5–10 years of data-engineering experience, can start right away, and have shipped at least one end-to-end Databricks project featuring Delta Lake and Unity Catalog, I’d love to review your profile and discuss how quickly we can get you onboarded.
Related categories:
SQL
Data Warehousing
Data Architecture
Data Governance
ETL
PySpark
Data Modeling
Big Data