I will build apache kafka, spark or flink data pipelines
About this Gig
I specialize in Apache Kafka, Spark, and Flink for high-throughput, real-time, and batch event processing.
Need to process events reliably at scale?
I design and build scalable event data pipelines for analytics, attribution, A/B testing, product events, and real-time data processing.
My experience includes systems processing 20B+ events per day, with a focus on high throughput, data quality, reliability, and maintainable architecture.
I can help you:
- Design scalable event ingestion and processing systems
- Build batch or real-time data pipelines
- Improve pipeline performance, cost, observability, and reliability
- Implement attribution and experimentation data flows
- Resolve data quality, duplication, ordering, and latency issues
You will receive practical, production-focused work aligned with your system scale and business goals.
Please contact me before ordering the Premium package to confirm scope, timeline, and required access.
FAQ
What types of pipelines can you build?
I can help with event ingestion, real-time processing, ETL/ELT, analytics, attribution, A/B testing, and data quality systems.
Which technologies do you work with?
I work primarily with Java, and Python. I can also work with your existing infrastructure, databases, queues, and cloud services.
Can you help with an existing pipeline?
Yes. I can review, diagnose, optimize, or extend an existing pipeline.
What do you need to begin?
Please share your current architecture, event volume, data sources, destination, technical stack, and main goal.
Can you handle very large event volumes?
Yes. I have experience with attribution and experimentation pipelines processing large amounts of data
Is implementation included?
Implementation is included only in the Premium package and depends on the agreed scope. Please contact me before ordering.
