Chatterbox

Chatterbox TTS

Resemble AI's state-of-the-art zero-shot voice cloning model with independent emotion control.

Registrer deg for 5000 tegn- grense

Bryt teksten i SSML- tagger for nøyaktig kontroll:

<speak><prosody rate="slow">Slow speech</prosody></speak>

Merker den valgte modellen forstår – klikk for å slippe en i teksten der den skjer:

Denne modellen leser ren tekst, så merker i teksten blir ignorert. Bytt til en ekspressiv modell som Orfeus eller Bark for å bruke tagger.

Definer selvvalgte uttaler (ord = uttale):

-12 +12
0.5x 2.0x
Fri for piper, VITS, MeloTTS
Her vises din genererte lyd. Velg en modell, skriv inn tekst og trykk Generer.
Lydgenerert vellykket
0:00
Last ned lyd Last ned.srt Lenke utløper om 24 timer
Fritt nivå: personlig bruk. Handelslisens fra $5/mo
Elsker TTS.ai? Fortell vennene dine!

Om Chatterbox

Chatterbox by Resemble AI is a leading open-source zero-shot voice cloning model that replicates a voice from a single audio sample, capturing not just timbre but speaking style and emotional nuance. Its distinctive feature is fine-grained emotion control that operates independently of the voice identity, so you can keep a cloned voice but shift its emotional intensity. Built around ResembleEnhance and flow matching, it targets professional-grade cloning for content creation, dubbing, and character voices. Released under the permissive MIT license, Chatterbox has become a popular foundation for derivative models — TTS.ai also runs a Saudi-Arabic fine-tune of it. It favors quality, with a modest per-request character limit.

Best for: Professional voice cloning with emotional control, content creation

Bla gjennom alle Chatterbox stemmer

Med et blikk

Utvikler
Resemble AI
Lisens
MIT
Nivå
premium
Hastighet
medium
Stemmekloning
Ja
Språk
English
Største antall tegn
300

Chatterbox stemmer

Default

English
Premie Neutral

Chatterbox TTS — OSS

A single audio sample is enough. Chatterbox is a zero-shot cloning model, so it captures a voice — including its style and emotional nuance — from one reference clip without any fine-tuning.

It offers fine-grained emotion control that works independently from the voice identity, letting you adjust the emotional tone of the output while keeping the same cloned voice.

Yes. Chatterbox is released by Resemble AI under the MIT license, which permits commercial use, and it serves as the base for several fine-tuned derivative models.
← Alle stemmer