Get contents from .*gz files over http using PySpark ETL Script, data pipe line
Budget: $30 – $50 USD
I have a file for an example test.csv.gz hosted on public server. The requirement is to have PySpark ETL script to fetch contents and send to elasticsearch. Implement sample test data pipeline with databricks on aws cloud.