Spark Script for Hive to S3 Data Migration

Job ID: 38029405

Budget: $30 – $250 AUD

Create a Spark script to transfer metastore data from Hive to S3

- Create a connection to Hive metastore
- Fetch schema.table definition for the database
- Create a connection to S3 bucket
- Create a new schema.table within S3 Hive metastore
- Transfer data from Hive metastore to S3
- Configure multiple schema.table creation based on config variables
- Create recursive data transfer based on difference in data

Skills and Experience:
- Proficiency in Spark and Hive
- Extensive experience with S3 buckets
- Understanding of data backup strategies

Project Details:
- The script needs to read the schema and perform metadata transfer for selected schema to s3 bucket.
- Only bid if you have work experience with spark, hive, s3.
- multiple schemas needs to be migrated.
- I have local instance of netapp s3 available and bucket created.
Related categories: Hive Spark Amazon S3 PySpark