Chatterbox Turbo

Chatterbox Turbo ТТС

A faster Chatterbox with sub-200ms latency and inline paralinguistic tags for laughs, coughs, and chuckles.

Жазылу 5000 таңба шегі

Мәтінді SSML тегтермен тасымалдау үшін нұсқау:

<speak><prosody rate="slow">Slow speech</prosody></speak>

Таңдалған үлгі түсінетін тегтер - біреуін мәтінге түсіру үшін түртіңіз:

Бұл модель қарапайым мәтін оқиды, сондықтан жол ішіндегі тегтер елеусіз қалып жатыр. Тегтерге негізделген эмоция үшін Orpheus не Bark сияқты өрнекті модельге ауысыңыз.

Өзінің дыбысын анықтау (сөз = дыбысы):

-12 +12
0.5x 2.0x
Piper, VITS, MeloTTS-пен тегінName
Бұл жерде құрылған аудио файлыңыз көрсетіледі. Үлгіні таңдап, мәтінін келтіріп, Құру дегенді басыңыз.
Аудио сәтті құрылды
0:00
Аудио жүктеп алу .srt жүктеп алу Сілтеменің мерзімі 24 сағаттан кейін аяқталады
Ұлттық құрама: Ұлттық құрама. Коммерциялық лицензия $5/ай-дан
Бұл дауысты өзіңіздің дауысыңыз қылып қойыңыз 30 секундта дауысты көшіріп алу
TTS.ai ұнады ма? Достарыңызға хабарлаңыз!

& Бұл туралы Chatterbox Turbo

Chatterbox Turbo is Resemble AI's 350M-parameter speed-focused upgrade to Chatterbox, reaching up to 6x real-time generation with sub-200-millisecond latency. It keeps the original's voice cloning while adding inline paralinguistic tags — you can drop [laugh], [cough], or [chuckle] directly into your text and have the model perform them. Every generation carries Perth watermarking for provenance tracking, a nod to responsible-AI deployment. The combination of low latency and expressive non-speech sounds makes it well suited to real-time voice agents and interactive characters. Like the original Chatterbox, it is MIT-licensed and English-focused.

Келесіге ең қолайлы: Real-time voice agents, expressive speech with natural sounds

Барлығын қарау Chatterbox Turbo дауыс

Бір қарапайым

Жасаушы
Resemble AI
Лицензия
MIT
Түр
standard
Жылдамдығы
fast
Дыбысын көшіру
Иә
Тілдер
English
Макс. таңбалар саны
1000

Chatterbox Turbo дауыс

Default

English
Әдетті Neutral

Chatterbox Turbo ТТС - ЖАҚСЫ СҰРАҚТАР

Turbo is a 350M-parameter model that runs at up to 6x real-time with sub-200ms latency, making it suitable for real-time voice agents where the original Chatterbox would be too slow.

They are inline cues like [laugh], [cough], and [chuckle] that you place directly in the text. The model performs the corresponding non-speech sounds, adding natural expressiveness.

Yes. All generated audio includes Perth watermarking, which supports provenance tracking and helps identify the output as AI-generated.
← Барлық дауыстар