Voices and engines
Voice profiles, the planned seven-engine TTS lineup (Qwen3-TTS, Chatterbox, Kokoro, and more), and the rules around voice cloning.
Voice profiles
Your voices live on your account as voice profiles — cloned voices and presets. Each profile carries a name, a language, the TTS engine it uses, and usage counts. When you generate speech (in the studio, through the MCP server, or through the API), you pick the profile to speak in.
TTS engines
Seven text-to-speech engines, including Qwen3-TTS, Chatterbox, and Kokoro, so you can pick the model that fits your quality, language, and speed needs on every generation. All inference runs on VoiceLabs' GPU servers — your browser never needs a GPU.
Engines change over time
The lineup above is what runs today. VoiceLabs posts real, dated release notes on the changelog whenever an engine is added, replaced or retired.
Voice cloning
Instant voice cloning from a short reference clip, in seconds. Two rules apply from day one:
- You may only clone a voice from a sample you are authorized to use.
- The Responsible Use policy applies to every generated voice.
AI-generated audio disclosure
Audio created with VoiceLabs is synthetic. Where the law or a platform requires it, disclose it as AI-generated. VoiceLabs marks exported audio machine-readably toward EU AI Act Article 50 (in force August 2, 2026).
Getting started
Create a VoiceLabs account in the browser, use your monthly free-generation allowance, and start the 7-day Pro trial.
Transcription and captures
Transcribe audio with VoiceLabs' hosted Whisper. Transcripts are stored as captures on your account, listable from the MCP server and the API.