I will do etl in pyspark

Pakistan

I speak English, German

6 orders completed

I am a data engineer with a passion for transforming raw data into meaningful insights. I have experience in designing, building, and maintaining scalable data pipelines using various tools and techno...
About this Gig

I will create an ETL (Extract. Transform, Load) pipeline for you in Pyspark. Depending on your preference I can create the ETL pipeline either in aws glue or databricks using Pyspark code.


Using Pyspark, I will get data from source whether it is some database, API, or data stored in AWS S3 bucket or Azure data lake.

After the data is loaded from source into staging area, I will transform it according to your business logic and load it into your target area.

The target area can be either data lake like S3 bucket, a data warehouse like redshift or some other place.

Other Data Engineering Services I Offer