I will design and build your etl elt data pipeline
Cloud Data Engineer
About this Gig
Struggling with messy, manual, or unreliable data workflows? I build clean, automated data
pipelines that just work reliably, on schedule, without babysitting.
I'm a Data Engineer working professionally on Spark-based ETL pipelines for financial data,
with hands-on experience in Python, PySpark, Airflow, PostgreSQL, and AWS.
WHAT I CAN BUILD FOR YOU:
ETL/ELT pipelines (API, database, or file-based sources)
Airflow DAGs for scheduled, automated workflows
Data validation & quality-check layers (null checks, record counts, schema checks)
PySpark pipelines for large dataset processing
PostgreSQL/MySQL data warehousing and reporting pipelines
Cleaning up and optimizing existing slow/unreliable pipelines
WHY WORK WITH ME:
- I write pipelines meant to run in production, not just once
- Clear communication I'll ask the right questions before writing a line of code
- Documented, handoff-ready code (README + comments), not a black box
Not sure exactly what you need? Message me with your data source and end goal I'll tell you
honestly what's realistic and how I'd approach it.
Expertise:
Big data
•
Data manipulation
•
ETL
•
Transformation
•
SQL
•
NoSQL
