I will build etl pipelines and databricks data engineering solutions
About this Gig
Data Engineer with 3+ years of hands-on experience building production-grade pipelines on Azure.
I specialize in designing and deploying scalable data workflows using Azure Data Factory, Databricks, PySpark, and Delta Lake from raw ingestion to analytics-ready Gold layer data models.
What I can build for you:
Metadata-driven ETL/ELT pipelines in Azure Data Factory
Medallion Architecture (Bronze Silver Gold) in Databricks
Real-time streaming pipelines using Azure Event Hubs (Kafka-compatible)
Star Schema / dimensional data warehouse design
SCD Type 1 & Type 2 change tracking with Databricks autoCDC
Complex SQL, PySpark transformations & stored procedures
REST/SOAP API integrations with Postman testing
Unity Catalog setup and data governance
Tech Stack:
Python · PySpark · SQL · Azure Data Factory · Azure Databricks · Azure Data Lake Gen2 · Delta Lake · Unity Catalog · Event Hubs · Power BI
Past highlights:
Publish layer pipelines at Oracle
Improved API efficiency by 50% via direct backend processing
Reduced Spark execution time by 30% through partitioning optimization
Built end-to-end ride-hailing real-time analytics platform on Azure
