Voice
Speech synthesis, zero-shot cloning, transcription, and dubbing — on your own models.
Synthesize
Clone a voice
Add a short, clean clip (5–10s). Zero-shot — no training step.
No cloning engine installed — install Chatterbox (MIT). Mock clones for UI testing only.
Transcribe
Audio → text (speech recognition).
Dub
Audio → transcribe → translate → re-voice (uses the selected voice).