Pyspark aws data engineer

Job ID: 37312338

Budget: $25 – $50 USD

I am looking for a Pyspark AWS data engineer who can help me with building and deploying ETL for machine learning models. Must initially pass a python online coding exam.

Tasks:
- Building ETL models using Pyspark and AWS
- Deploying the models on AWS infrastructure
- use terraform, spin up etl clusters, understand basic data related aws cloud tools, infrastructure and security. This is NOT a devops position but you should be able to get around and use data engineering related aws tools.

Infrastructure:
- The project requires migrating within aws to a new infrastructure

Involvement:
- partially involved in the project at half time 3-5 hours a day on a consistent reliable time of your choosing.

Ideal skills and experience:
- Strong experience in data engineering with Pyspark and AWS,
- Knowledge of etl and cluster deployment on AWS
- Ability to modify basic infrastructure on AWS

Note: Please only apply if you have experience with building and deploying etl and data models using Pyspark and AWS. Strong enough python coding to pass an online exam once.
Related categories: Python Amazon Web Services Spark ETL Terraform