Voices and engines
Voice profiles, the planned seven-engine TTS lineup (Qwen3-TTS, Chatterbox, Kokoro, and more), and the rules around voice cloning.
Voice profiles
Your voices live on your account as voice profiles — cloned voices and presets. Each profile carries a name, a language, the TTS engine it uses, and usage counts. When you generate speech (through the MCP server or the API today, and the studio at launch), you pick the profile to speak in.
TTS engines
Seven text-to-speech engines are planned at launch, including Qwen3-TTS, Chatterbox, and Kokoro, so you can pick the model that fits your quality, language, and speed needs. All inference runs on VoiceLabs' GPU servers — your browser never needs a GPU.
Pre-launch status
The engine lineup and instant voice cloning are planned at launch and are described here as such — VoiceLabs posts real, dated release notes on the changelog as features ship.
Voice cloning
Instant voice cloning from a short reference clip is planned at launch. Two rules apply from day one:
- You may only clone a voice from a sample you are authorized to use.
- The Responsible Use policy applies to every generated voice.
AI-generated audio disclosure
Audio created with VoiceLabs is synthetic. Where the law or a platform requires it, disclose it as AI-generated. VoiceLabs marks exported audio machine-readably toward EU AI Act Article 50 (in force August 2, 2026).
Getting started
Create a VoiceLabs account in the browser, use your monthly free-generation allowance, and start the 7-day Pro trial.
Transcription and captures
Transcribe audio with VoiceLabs' hosted Whisper. Transcripts are stored as captures on your account, listable from the MCP server and the API.