VoiceLabs

Voices and engines

Voice profiles, the planned seven-engine TTS lineup (Qwen3-TTS, Chatterbox, Kokoro, and more), and the rules around voice cloning.

Voice profiles

Your voices live on your account as voice profiles — cloned voices and presets. Each profile carries a name, a language, the TTS engine it uses, and usage counts. When you generate speech (through the MCP server or the API today, and the studio at launch), you pick the profile to speak in.

TTS engines

Seven text-to-speech engines are planned at launch, including Qwen3-TTS, Chatterbox, and Kokoro, so you can pick the model that fits your quality, language, and speed needs. All inference runs on VoiceLabs' GPU servers — your browser never needs a GPU.

Pre-launch status

The engine lineup and instant voice cloning are planned at launch and are described here as such — VoiceLabs posts real, dated release notes on the changelog as features ship.

Voice cloning

Instant voice cloning from a short reference clip is planned at launch. Two rules apply from day one:

  • You may only clone a voice from a sample you are authorized to use.
  • The Responsible Use policy applies to every generated voice.

AI-generated audio disclosure

Audio created with VoiceLabs is synthetic. Where the law or a platform requires it, disclose it as AI-generated. VoiceLabs marks exported audio machine-readably toward EU AI Act Article 50 (in force August 2, 2026).

On this page