Chatterbox Turbo

Chatterbox Turbo ТТС

A faster Chatterbox with sub-200ms latency and inline paralinguistic tags for laughs, coughs, and chuckles.

Упиши се за ограничење 5.000 знакова

Умотајте текст у ССМЛ ознаке за прецизну контролу:

<speak><prosody rate="slow">Slow speech</prosody></speak>

Означе изабрани модел разуме — кликните да га испустите у текст где се то дешава:

Овај модел чита обичан текст, па се ознаке за успостављене на линији игноришу. За емоције засноване на ознакама пребаците на изражавајући модел попут Орфеја или Барка.

Дефинишите посебне изговоре (слов = изговор):

-12 +12
0.5x 2.0x
Слободна са Пајпер, Витс, Мелоттс
Овд› је ће се појавити генерисани звук. Изаберите модел, унесите текст и кликните на Генериши.
аудио генерисано усп› јешно
0:00
Преузми аудио Преузми.рт Веза истекава за 24х
Слободни ниво: лично коришћење. Комерцијална дозвола од 5 долара/мо
Нека ово буде твој глас Клонирај глас за 30 секунди
Љубав ТТС.аи?

О Chatterbox Turbo

Chatterbox Turbo is Resemble AI's 350M-parameter speed-focused upgrade to Chatterbox, reaching up to 6x real-time generation with sub-200-millisecond latency. It keeps the original's voice cloning while adding inline paralinguistic tags — you can drop [laugh], [cough], or [chuckle] directly into your text and have the model perform them. Every generation carries Perth watermarking for provenance tracking, a nod to responsible-AI deployment. The combination of low latency and expressive non-speech sounds makes it well suited to real-time voice agents and interactive characters. Like the original Chatterbox, it is MIT-licensed and English-focused.

Најбоље за: Real-time voice agents, expressive speech with natural sounds

Прегледај све Chatterbox Turbo гласови

На један поглед

Програмер
Resemble AI
Лиценца
MIT
Низ
standard
Брзина
fast
Гласово клонирање
Да.
језици
English
Макс. знакова
1000

Chatterbox Turbo гласови

Default

English
стандардни Neutral

Chatterbox Turbo ТТС — честа питања

Turbo is a 350M-parameter model that runs at up to 6x real-time with sub-200ms latency, making it suitable for real-time voice agents where the original Chatterbox would be too slow.

They are inline cues like [laugh], [cough], and [chuckle] that you place directly in the text. The model performs the corresponding non-speech sounds, adding natural expressiveness.

Yes. All generated audio includes Perth watermarking, which supports provenance tracking and helps identify the output as AI-generated.
← Сви гласови