I will deploy comfyui as a scalable runpod serverless API

A
avivk5498
A
avivk5498
Aviv K.

About this gig

Your ComfyUI workflow works locally. Now you need it to behave like a product.

I will deploy your existing image or video workflow to Runpod as a production API your application can call.


This is more than a basic pod setup. I build the surrounding backend needed to run GPU inference reliably, securely, and at scale.


Depending on your package, the system can include:

  • Dockerized ComfyUI deployment
  • Authenticated API with dynamic workflow inputs
  • Queued execution and autoscaling GPU workers
  • Private output storage and webhooks
  • Retries, timeouts, logs, and monitoring
  • Cold-start, VRAM, and GPU cost optimization
  • Source code, API documentation, and technical handoff


I work with Wan, LTX, MiniMax, Qwen, Z-Image, Krea, Flux, and Stable Diffusion workflows.

One Runpod backend I built processed more than 45,000 jobs at 99.9% uptime while reducing GPU costs by roughly 60%.

Before ordering, send me your workflow JSON, required models, expected traffic, and API inputs. I will confirm the appropriate package and identify any workflow changes required before deployment.

Get to know Aviv K.

Aviv K.

Senior AI Engineer, Runpod and ComfyUI, Solutions Architect

  • FromIsrael
  • Member sinceAug 2026
  • Languages

    English, Hebrew
I'm a senior software engineer and solutions architect who turns AI ideas into systems teams can run. I design and build governed AI agents, AWS serverless workflows, Runpod and ComfyUI backends, image and video APIs, and LoRA training pipelines. Before specializing in AI, I handled more than 250 production incidents. One generation backend I built processed 45,000+ jobs at 99.9% uptime while reducing GPU costs by about 60%. I work from diagnosis and architecture through deployment, monitoring, and handoff. If you need technical direction and an engineer who can build the system, let's talk.

My Portfolio