I will build python etl pipelines for your data warehouse
Product Engineering Intern
About this Gig
I'm a Computer Science engineering student with hands-on experience building ETL pipelines and data engineering solutions using Python, SQL, and Databricks. Through real internship work, I've built backend services and REST APIs that process large-scale data (120M+ records), integrated Databricks, Unity Catalog, and Delta Tables into production workflows, and implemented data validation and cleanup pipelines for engineering teams.
What I offer:
ETL pipeline design & development (extract, transform, load)
Data source connectivity (databases, APIs, files)
Data cleaning, formatting, and validation
Databricks & Delta Table workflows
API integration for automated data flows
Clean, documented, reusable code
I focus on writing pipelines that are reliable, well-structured, and easy to maintain not just quick scripts.
I communicate clearly throughout the project and deliver on time.
Expertise:
API integration
•
Big data
•
Data manipulation
•
ETL
•
SQL
•
NoSQL
Technology:
Apache Spark
•
Java
•
Python
•
R
•
SQL
•
Databricks
My Portfolio
FAQ
Do you work with data warehouses other than Databricks?
Databricks is my primary expertise. For other platforms, message me first so I can confirm fit. But I can work, If once I get comfortable with it.
Will I get the source code?
Yes, included in the Premium package (or as an add-on).
Can you handle multiple data sources?
Yes — Standard and Premium packages support multiple sources.
What info do you need to start?
Your data source(s), target output/warehouse, and any specific transformation logic.
