I will build your music ai app, voice assistant, or audio pipeline

C
chen777777
C
chen777777
Chen W.

About this gig

Need a music, voice, or audio AI feature that works beyond a demo?


I build Python-based audio AI products for creators, startups, and research teams. I hold an MSc in Sound and Music Computing from Queen Mary University of London and have 4 years of AI R&D experience.


Relevant work:

- EmoHeal: fine-grained emotion analysis and therapeutic music retrieval/generation

- VoxSketch: spoken brief to editable Audiotool song structure through NEXUS


I can deliver:

- generative or adaptive music workflows

- speech-to-text, text-to-speech, and voice interfaces

- audio classification, feature extraction, or emotion analysis

- voice-enabled AI agents

- FastAPI APIs, Docker, source code, tests, and handoff notes


Basic covers one focused feature. Standard covers an integrated audio pipeline. Premium covers a custom multi-stage music or voice AI system.


Please message me before ordering with your sample inputs, target output, platform, and deadline. Voice cloning, commercial music rights, real-time streaming, model training, hosting, and third-party fees require scope and licensing review.

Get to know Chen W.

Chen W.

Music, Voice AI and Agent Developer

  • FromChina
  • Member sinceJul 2026
  • Avg. response time1 hour
  • Languages

    Chinese, English
I build music, voice, and agent AI products from prototype to tested delivery. I hold an MSc in Sound and Music Computing from QMUL and have 4 years of AI R&D experience. My work includes EmoHeal—emotion analysis plus therapeutic music retrieval/generation—and VoxSketch, a voice-to-editable-song workflow using Audiotool NEXUS. I deliver generative music workflows, STT/TTS, audio analysis, voice agents, RAG, FastAPI, Docker, source code, tests, and documentation. Expect clear scope, honest technical boundaries, milestone delivery, and a system you can keep building after handoff.