Chatterbox Turbo

Chatterbox Turbo TTS

A faster Chatterbox with sub-200ms latency and inline paralinguistic tags for laughs, coughs, and chuckles.

Bhalisa Uluhlu lwezinto zobumnini Zolwaleko...

Ulawulo oluchanekileyo:

<speak><prosody rate="slow">Slow speech</prosody></speak>

Ii-tags imodeli ekhethiweyo iqonda - nqakraza ukushiya enye kumbhalo wakho apho isenza khona:

Le modeli ifunda umbhalo oqhelekileyo, ngoko ke i-inline tags ilahleka. Uphawu olusekelwe kwi-emotions, tshintshela kwimodeli ebonisa umbono njenge-Orpheus okanye i-Bark.

Chaza ubeko lwephepha

-12 +12
0.5x 2.0x
Ikhululekile nge Piper, VITS, MeloTTS
Isandi sakho esivelisweyo siza kuvela apha. Khetha imodeli, ngenisa umbhalo, kwaye unqakraze Yenza.
Isandi Sizaliswe Ngempumelelo
0:00
Layisha ezantsi Layisha ezantsi Ikhonkco liphelelwe lixesha kwiyure ezi-24
Inqanaba elikhululekileyo: ukusetyenziswa komuntu siqu. Ilayisensi yezorhwebo ukusuka kwi- $5/inyanga
Uthando TTS.ai? Nceda utshele abalandeli bakho!

I-About Chatterbox Turbo

Chatterbox Turbo is Resemble AI's 350M-parameter speed-focused upgrade to Chatterbox, reaching up to 6x real-time generation with sub-200-millisecond latency. It keeps the original's voice cloning while adding inline paralinguistic tags — you can drop [laugh], [cough], or [chuckle] directly into your text and have the model perform them. Every generation carries Perth watermarking for provenance tracking, a nod to responsible-AI deployment. The combination of low latency and expressive non-speech sounds makes it well suited to real-time voice agents and interactive characters. Like the original Chatterbox, it is MIT-licensed and English-focused.

Elungileyo: Real-time voice agents, expressive speech with natural sounds

Khangela konke Chatterbox Turbo iilizwi

Kwingxelo

Umbhekisi phambili
Resemble AI
Ilayisensi
MIT
I-Tier
standard
Isantya
fast
Ukuphinda usebenzise ilizwi
Ewe
Iilwimi
English
Ubukhulu bamagama
1000

Chatterbox Turbo iilizwi

Default

English
Emiselweyo Neutral

Chatterbox Turbo TTS - Imibuzo ebuzwa rhoqo

Turbo is a 350M-parameter model that runs at up to 6x real-time with sub-200ms latency, making it suitable for real-time voice agents where the original Chatterbox would be too slow.

They are inline cues like [laugh], [cough], and [chuckle] that you place directly in the text. The model performs the corresponding non-speech sounds, adding natural expressiveness.

Yes. All generated audio includes Perth watermarking, which supports provenance tracking and helps identify the output as AI-generated.
← Zonke iingoma