I deploy AI systems that work in production. I've gone deep on LLM infrastructure, resolving GPU memory and KV cache issues at scale, and integrating locally-hosted models into production backends. My background is full-stack engineering with a solid DevOps foundation.... Read more