Pyspark & EMR developer needed

Job ID: 35338957

Budget: ₹75,000 – ₹150,000 INR

Total Experience : 4+ years to 7 years
Designation : Sr. Data Engineer
Mandatory skills : Pyspark & EMR
Location : Pune /Remote



Job Description -
1) Hands-on experience with Python, Spark, EMR
2) Proficient understanding of distributed computing principles
3) Proficiency with Data Processing: HDFS, Hive, Spark, Scala/Python
4) Independent thinker, willing to engage, challenge and learn new technologies.
5) Understanding of the benefits of data warehousing, data architecture, data quality processes, data warehousing design, and implementation,
6) Table structure, fact and dimension tables, logical and physical database design, data modeling, reporting process metadata, and ETL processes.

Requirements --

1) Client-facing skills: Solid experience working with clients directly, to be able to build trusted relationships with stakeholders.
2) In-depth understanding of Data Warehouse, ETL concept and modeling structure principles
3) Expertise in AWS cloud native services
4) Hand-on experience in developing data processing task using Spark on cloud native services like Glue/EMR.
5) Good to have experience Snowflake SQL queries against Snowflake Developing scripts using java scripts to do Extract, Load, and Transform data
6) Good to have experience with Snowflake utilities such as SnowSQL, SnowPipe, Python, Tasks, Streams, Time travel, Optimizer, Metadata Manager, data sharing, and stored procedures.
7) Excellent verbal and written communications skills
8) Ability to collaborate effectively across global teams

Please apply if you're an independent freelancer only & available for a full-time contractual job.

*Agencies please do not apply*
Related categories: Python Hadoop Hive Spark PySpark