Data Operations Engineer
Budget: ₹750 – ₹1,250 INR
● Experience wrangling terabytes of big, complicated, imperfect data
● Extensive experience designing and implementing ETL pipelines
● Experience with ETL job automation through Airflow pipelines
● Experience building and operationalizing large-scale enterprise data solutions, Data Lakes and applications using one or more of AWS data and analytics services
● Experience with AWS products (Redshift, EC2, EMR, S3, IAM, RDS, CloudWatch etc) and Databricks
● Bachelor's degree in Computer Science or a related field (or 4 additional years of relevant work experience)
● A strong understanding of data structures, algorithms, and effective software design
Significant development experience with a major modern language (e.g. Java, Scala, Python, Ruby, Spark etc.)
● Significant experience working with structured and unstructured data at scale and comfort with a variety of different stores (key-value, document, columnar, etc.) as well as traditional RDBMSes and data warehouses
● Experience with or interest in AWS Glue, Redshift Spectrum and any other tools that enable data querying at scale
● Exposure to visualization tools - Tableau, Google Analytics, JIRA, MS Project
● Conduct code review with architecture team to ensure standards best practices are followed
● Develop automated code deploys with GIT, AWS Lambda and AWS Batch services
● Ability to multi-task and manage multiple environments
● Provide on-call support and remote troubleshooting
● Must work well in an agile, collaborative team environment
● Extensive experience designing and implementing ETL pipelines
● Experience with ETL job automation through Airflow pipelines
● Experience building and operationalizing large-scale enterprise data solutions, Data Lakes and applications using one or more of AWS data and analytics services
● Experience with AWS products (Redshift, EC2, EMR, S3, IAM, RDS, CloudWatch etc) and Databricks
● Bachelor's degree in Computer Science or a related field (or 4 additional years of relevant work experience)
● A strong understanding of data structures, algorithms, and effective software design
Significant development experience with a major modern language (e.g. Java, Scala, Python, Ruby, Spark etc.)
● Significant experience working with structured and unstructured data at scale and comfort with a variety of different stores (key-value, document, columnar, etc.) as well as traditional RDBMSes and data warehouses
● Experience with or interest in AWS Glue, Redshift Spectrum and any other tools that enable data querying at scale
● Exposure to visualization tools - Tableau, Google Analytics, JIRA, MS Project
● Conduct code review with architecture team to ensure standards best practices are followed
● Develop automated code deploys with GIT, AWS Lambda and AWS Batch services
● Ability to multi-task and manage multiple environments
● Provide on-call support and remote troubleshooting
● Must work well in an agile, collaborative team environment