Chatterbox Turbo

Chatterbox Turbo ТТС

A faster Chatterbox with sub-200ms latency and inline paralinguistic tags for laughs, coughs, and chuckles.

Запиши се. за 5000 ограничаване на знака

Опаковка на вашия текст в SSML тагове за точен контрол:

<speak><prosody rate="slow">Slow speech</prosody></speak>

Етикети избрания модел разбира — кликнете за да пуснете един в вашия текст, където се случва:

Този модел чете обикновен текст, така че в линия тагове се пренебрегват. За емоции, базирани на таг, превключете на експресивен модел като Orpheus или Bark.

Определяне на своите изговори (слово = произношение):

-12 +12
0.5x 2.0x
Без пари с Пайпър, ВИТС, Мелотс
Тук ще се появи генерираното ви аудио. Изберете модел, въведете текст и кликнете върху Генериране.
Аудио генерирано успешно
0:00
Изтегляне на аудиото Сваляне.rt. Връзката изтича след 24 часа
Безплатен ступенят: лично използване. Търговска лиценз от $5/мо
Обичай ТТСай, кажи на приятелите си!

За Chatterbox Turbo

Chatterbox Turbo is Resemble AI's 350M-parameter speed-focused upgrade to Chatterbox, reaching up to 6x real-time generation with sub-200-millisecond latency. It keeps the original's voice cloning while adding inline paralinguistic tags — you can drop [laugh], [cough], or [chuckle] directly into your text and have the model perform them. Every generation carries Perth watermarking for provenance tracking, a nod to responsible-AI deployment. The combination of low latency and expressive non-speech sounds makes it well suited to real-time voice agents and interactive characters. Like the original Chatterbox, it is MIT-licensed and English-focused.

Най-добро за: Real-time voice agents, expressive speech with natural sounds

Преглед на всички Chatterbox Turbo гласове

На един поглед.

Разработчик
Resemble AI
Лиценз
MIT
Ниво на равнището
standard
Скорост
fast
Гласово клониране
Да.
Езици
English
Макс. символи
1000

Chatterbox Turbo гласове

Default

English
Стандартен Neutral

Chatterbox Turbo TTS — Често задавани въпроси

Turbo is a 350M-parameter model that runs at up to 6x real-time with sub-200ms latency, making it suitable for real-time voice agents where the original Chatterbox would be too slow.

They are inline cues like [laugh], [cough], and [chuckle] that you place directly in the text. The model performs the corresponding non-speech sounds, adding natural expressiveness.

Yes. All generated audio includes Perth watermarking, which supports provenance tracking and helps identify the output as AI-generated.
← Всички гласове.