Chatterbox

Chatterbox TTS

Resemble AI's state-of-the-art zero-shot voice cloning model with independent emotion control.

Bhalisa Uluhlu lwezinto zobumnini Zolwaleko...

Ulawulo oluchanekileyo:

<speak><prosody rate="slow">Slow speech</prosody></speak>

Ii-tags imodeli ekhethiweyo iqonda - nqakraza ukushiya enye kumbhalo wakho apho isenza khona:

Le modeli ifunda umbhalo oqhelekileyo, ngoko ke i-inline tags ilahleka. Uphawu olusekelwe kwi-emotions, tshintshela kwimodeli ebonisa umbono njenge-Orpheus okanye i-Bark.

Chaza ubeko lwephepha

-12 +12
0.5x 2.0x
Ikhululekile nge Piper, VITS, MeloTTS
Isandi sakho esivelisweyo siza kuvela apha. Khetha imodeli, ngenisa umbhalo, kwaye unqakraze Yenza.
Isandi Sizaliswe Ngempumelelo
0:00
Layisha ezantsi Layisha ezantsi Ikhonkco liphelelwe lixesha kwiyure ezi-24
Inqanaba elikhululekileyo: ukusetyenziswa komuntu siqu. Ilayisensi yezorhwebo ukusuka kwi- $5/inyanga
Uthando TTS.ai? Nceda utshele abalandeli bakho!

I-About Chatterbox

Chatterbox by Resemble AI is a leading open-source zero-shot voice cloning model that replicates a voice from a single audio sample, capturing not just timbre but speaking style and emotional nuance. Its distinctive feature is fine-grained emotion control that operates independently of the voice identity, so you can keep a cloned voice but shift its emotional intensity. Built around ResembleEnhance and flow matching, it targets professional-grade cloning for content creation, dubbing, and character voices. Released under the permissive MIT license, Chatterbox has become a popular foundation for derivative models — TTS.ai also runs a Saudi-Arabic fine-tune of it. It favors quality, with a modest per-request character limit.

Elungileyo: Professional voice cloning with emotional control, content creation

Khangela konke Chatterbox iilizwi

Kwingxelo

Umbhekisi phambili
Resemble AI
Ilayisensi
MIT
I-Tier
premium
Isantya
medium
Ukuphinda usebenzise ilizwi
Ewe
Iilwimi
English
Ubukhulu bamagama
300

Chatterbox iilizwi

Default

English
Ixabiso eliphezulu Neutral

Chatterbox TTS - Imibuzo ebuzwa rhoqo

A single audio sample is enough. Chatterbox is a zero-shot cloning model, so it captures a voice — including its style and emotional nuance — from one reference clip without any fine-tuning.

It offers fine-grained emotion control that works independently from the voice identity, letting you adjust the emotional tone of the output while keeping the same cloned voice.

Yes. Chatterbox is released by Resemble AI under the MIT license, which permits commercial use, and it serves as the base for several fine-tuned derivative models.
← Zonke iingoma