Chatterbox

Chatterbox TTS

Resemble AI's state-of-the-art zero-shot voice cloning model with independent emotion control.

Inscrever-se para o limite de 5000 caracteres

Envolva o seu texto em tags SSML para controle preciso:

<speak><prosody rate="slow">Slow speech</prosody></speak>

Etiquetas o modelo selecionado entende — clique para soltar um para o seu texto onde acontece:

Este modelo lê texto simples, por isso as etiquetas inline são ignoradas. Para emoção baseada em tags, mude para um modelo expressivo como Orpheus ou Bark.

Definir pronúncias personalizadas (palavra = pronúncia):

-12 +12
0.5x 2.0x
Grátis com Piper, VITS, MeloTTS
Seu áudio gerado aparecerá aqui. Escolha um modelo, introduza texto e clique em Gerar.
O áudio gerado com sucesso
0:00
Baixe áudio Baixar.srt A ligação expira em 24h
Gratuito nível: uso pessoal. Licença comercial a partir de $5/mo
Gosta do TTS.ai? Conte aos seus amigos!

Sobre Chatterbox

Chatterbox by Resemble AI is a leading open-source zero-shot voice cloning model that replicates a voice from a single audio sample, capturing not just timbre but speaking style and emotional nuance. Its distinctive feature is fine-grained emotion control that operates independently of the voice identity, so you can keep a cloned voice but shift its emotional intensity. Built around ResembleEnhance and flow matching, it targets professional-grade cloning for content creation, dubbing, and character voices. Released under the permissive MIT license, Chatterbox has become a popular foundation for derivative models — TTS.ai also runs a Saudi-Arabic fine-tune of it. It favors quality, with a modest per-request character limit.

Melhor para: Professional voice cloning with emotional control, content creation

Procurar todos Chatterbox vozes

De uma olhada

Desenvolvedor
Resemble AI
Licença
MIT
Tier
premium
Velocidade
medium
Clonagem de voz
Sim
Línguas
English
Número máximo de caracteres
300

Chatterbox vozes

Default

English
Premium Neutral

Chatterbox TTS — FAQ

A single audio sample is enough. Chatterbox is a zero-shot cloning model, so it captures a voice — including its style and emotional nuance — from one reference clip without any fine-tuning.

It offers fine-grained emotion control that works independently from the voice identity, letting you adjust the emotional tone of the output while keeping the same cloned voice.

Yes. Chatterbox is released by Resemble AI under the MIT license, which permits commercial use, and it serves as the base for several fine-tuned derivative models.
← Todas as vozes