
Ashish Kumar
Data Engineer
Skills

See my services


Work experience
Staff Engineer II
Biome Analysis • Full-time
Oct 2024 - Present • 1 yr 10 mos
I spearheaded the development of automated data ingestion processes that transformed how our team handled complex data mappings and transformations, cutting manual effort in half through intelligent automation at the loading phase. By developing Python scripts to automate configuration file generation, I eliminated 4-5 hours of setup time per configuration, allowing the team to focus on higher-value work. Working closely with cross-functional teams, I helped optimize our data processing pipelines for better performance and reliability. To ensure data quality at scale, I implemented comprehensive validation checks that maintained a 99.9% accuracy rate across all datasets processed daily, significantly reducing data inconsistencies and building trust in our data infrastructure.
Biome Analytics
Full-time • 3 yrs
Staff Engineer I
Oct 2023 - Oct 2024 • 1 yr
I designed and implemented scalable data ingestion pipelines that transformed our system's performance, reducing processing time by 50% while seamlessly scaling to handle over 10 GB of data. Through careful analysis of our data ingestion lifecycle, I identified and resolved critical issues in legacy code that had been causing data errors, ultimately achieving a 99.9% accuracy rate across all datasets. Collaborating with cross-functional teams, I drove the implementation of new features that enhanced system functionality and delivered an additional 20% reduction in processing time, directly boosting team productivity. My commitment to data quality extended to establishing continuous monitoring processes that proactively identified and resolved inconsistencies, maintaining 99.9% data integrity and contributing to 100% system uptime—ensuring our data infrastructure remained reliable and trustworthy for all stakeholders.
Data Engineer
Oct 2021 - Oct 2023 • 2 yrs
I manage end-to-end data ingestion processes and troubleshoot Data Assimilation code to ensure reliable data flow across our infrastructure. Leveraging Azure Data Factory pipelines, I orchestrate the seamless movement of approximately 500 files between repositories while implementing rigorous data validation procedures to maintain integrity and quality standards. On a daily basis, I supervise 3-6 ETL tasks and proactively address an average of 3 critical errors within 4-5 hours, consistently achieving a 99% processing success rate through rapid issue resolution and careful monitoring. Throughout these workflows, I utilize SQL and Python to query, manipulate, and process data efficiently, ensuring our data operations run smoothly and meet the demands of our business stakeholders.