Chatterbox

Chatterbox TTS

Resemble AI's state-of-the-art zero-shot voice cloning model with independent emotion control.

Ṣẹ̀dà fun àwọn àmì-àṣírí 5,000

Fi àkọlé rẹ pamọ́ sí àwọn àmì-ìwé SSML fún ìdáràn:

<speak><prosody rate="slow">Slow speech</prosody></speak>

Àwọn Àmì-ìwé tí àwọn ìṣàmúlò-ètò tí a yàn gbọ́ - tẹ̀ láti fi ọkan sínú àkọ́lé rẹ̀ nínú àwọn ààyè-iṣẹ́ tí o bá jẹ́:

Àwọn àwọn àkọlé àwọn ààyè-iṣẹ́ àwọn àwọn àmì-ìwé àwọn àmì-ìwé àwọn àwọn àmì-ìwé àwọn à

Àwọn àwọn ìṣàfarawé àwọn àwọn ìṣàfarawé àwọn (ọrọ = ìṣàfàlì):

-12 +12
0.5x 2.0x
Free pẹlu Piper, VITS, MeloTTS
Àwọn àwòrán tí o ti ṣẹ̀dà tí o bá han níbẹ̀. Yan àwọn àwòrán, tẹ̀lẹ̀ àkọlé, ki o si tẹ̀ Ṣẹ̀dà.
Àwọn àwọn àwòrán tí a ṣẹ̀dà
0:00
Ṣàfikún Àwọn Àmì-ìwé Ṣàfikún.srt Líǹkì náà kù nínú 24h
Ìjádé ọ̀fẹ́: ìlòjútó ara ẹni. Lisensi Iṣowo ori lati $5/mo
O fẹ́ TTS.ai? Fì sọ̀kalẹ̀ fún àwọn ọrẹ̀ rẹ̀!

Ààyè-iṣẹ́ Chatterbox

Chatterbox by Resemble AI is a leading open-source zero-shot voice cloning model that replicates a voice from a single audio sample, capturing not just timbre but speaking style and emotional nuance. Its distinctive feature is fine-grained emotion control that operates independently of the voice identity, so you can keep a cloned voice but shift its emotional intensity. Built around ResembleEnhance and flow matching, it targets professional-grade cloning for content creation, dubbing, and character voices. Released under the permissive MIT license, Chatterbox has become a popular foundation for derivative models — TTS.ai also runs a Saudi-Arabic fine-tune of it. It favors quality, with a modest per-request character limit.

Tí o dara jù fún: Professional voice cloning with emotional control, content creation

Wá Gbogbo àwòrán Chatterbox Àwọn àwòrán

Nínú àwọn ìṣàfarawé

Àwọn Àkọlé
Resemble AI
Àwọn Ààyè-iṣẹ́
MIT
Àwọn àwọn ààyè-iṣẹ́
premium
Ìjánu-ìṣàmúlò-ètò
medium
Ìṣàfarawé àwọn àmì-ìwé
Yà
Àwọn
English
Àwọn àyọkà ìpele
300

Chatterbox Àwọn àwòrán

Default

English
Àwọn ìṣàmúlò-ètò Neutral

Chatterbox Àwọn Àtòjọ-ẹ̀yàn

A single audio sample is enough. Chatterbox is a zero-shot cloning model, so it captures a voice — including its style and emotional nuance — from one reference clip without any fine-tuning.

It offers fine-grained emotion control that works independently from the voice identity, letting you adjust the emotional tone of the output while keeping the same cloned voice.

Yes. Chatterbox is released by Resemble AI under the MIT license, which permits commercial use, and it serves as the base for several fine-tuned derivative models.
← Gbogbo àwọn ìrànwọ́