I will build robust and scalable data pipelines using pyspark, databricks, and airflow

India

I speak English

Certified Data Engineer

Results-driven Data Engineer with 4.5 years of experience in designing, developing, testing, supporting and optimizing scalable ETL/ELT pipelines using PySpark, Databricks, SQL, Python, Apache Spark a...
About this Gig

Need scalable, production-ready data pipelines? I build high-performance data architectures using Databricks, PySpark, SQL, Apache Airflow, and Apache Kafka. From simple file ingestion to complex real-time streaming, I deliver optimized, cost-efficient cloud solutions.


What I Offer:

  • End-to-End ETL/ELT: Secure data pipelines using PySpark and SQL.
  • Databricks Optimization: Delta Lake, Delta Live Tables (DLT), and Unity Catalog setup.
  • Orchestration: Automated Apache Airflow DAGs with error handling.
  • Real-Time Streaming: Low-latency event streaming via Apache Kafka.
  • Performance Tuning: Slashing your cloud compute costs with optimized queries.


Why Choose Me?

  • Production-grade, clean, and reusable code.
  • Enterprise cloud experience (AWS / Azure).
  • Free architectural documentation included.



Destination Platform:

Snowflake

Databricks Lakehouse

Teradata

Tools & Platforms:

Fivetran

Kafka Connect

Microsoft SSIS

Other