I will build etl data pipelines with pyspark and databricks

Nepal

I speak English, Nepali, Hindi

Data Engineer

I'm a Data & Backend Engineer with 5+ years of experience building scalable ETL pipelines, cloud data platforms, and robust backend systems. My expertise spans Python, PySpark, SQL, AWS, GCP, Databric...
About this Gig

Hello, I am Arbind Sah, and I can help you build scalable ETL pipelines, process Big Data, or optimize cloud-based systems.

I specialize in Python, SQL, and cloud-native tools to streamline data integration, automation, and analytics for businesses of all sizes.


What I Can help you with:

-ETL & Data Pipelines Design and optimize ETL/ELT workflows using PySpark and Databricks for reliable, production-ready data transformation.Solutions Utilize Redshift, BigQuery, Spark, and Databricks to handle massive datasets efficiently.

-Cloud Infrastructure: Build and manage infrastructure on AWS (EC2, RDS, S3, Lambda) and GCP, including containerized deployments with Docker.

-Database Optimization & Security Improve query performance, implement robust security, and ensure compliance with industry standards.

-Data Migration & Schema Reconciliation Migrate and consolidate data across PostgreSQL, MySQL and cloud platforms, handling type mismatches, null values, and complex schema mapping.


With a strong focus on cloud computing, data automation, and analytics, I ensure scalable, secure, and high-performance data solutions.

Connect with me to optimize your data workflows for maximum efficiency!

Destination Platform:

Google BigQuery

Databricks Lakehouse

Tools & Platforms:

Airbyte

AWS Glue DataBrew

My Portfolio