Chatterbox Turbo

Chatterbox Turbo TTS

A faster Chatterbox with sub-200ms latency and inline paralinguistic tags for laughs, coughs, and chuckles.

Enskri Limit pou 5,000 karaktè

Wrap ou tèks nan SSML tags pou presizyon kontwòl:

<speak><prosody rate="slow">Slow speech</prosody></speak>

Tags ke modèl la chwazi konprann — klike pou mete yon nan tèks ou kote li rive:

Modèl sa a li tèks senp, se poutèt sa atik ki nan liy yo pa pran an kont. Pou efè ki baze sou atik, chanje pou yon modèl ekspresyon tankou Orpheus oswa Bark.

Define prononciations Custom (mot = prononciation):

-12 +12
0.5x 2.0x
Gratis ak Piper, VITS, MeloTTS
Son ou kreye a ap parèt isit la. Chwazi yon modèl, antre tèks la, epi klike Kreye.
Audio Generated Successfully
0:00
Telechaje son Telechaje.srt Link expires in 24h
Free tier: itilize pèsonèl. Lisans Komèsyal soti nan $5/mo
Love TTS.ai? Di zanmi ou yo!

Atik Chatterbox Turbo

Chatterbox Turbo is Resemble AI's 350M-parameter speed-focused upgrade to Chatterbox, reaching up to 6x real-time generation with sub-200-millisecond latency. It keeps the original's voice cloning while adding inline paralinguistic tags — you can drop [laugh], [cough], or [chuckle] directly into your text and have the model perform them. Every generation carries Perth watermarking for provenance tracking, a nod to responsible-AI deployment. The combination of low latency and expressive non-speech sounds makes it well suited to real-time voice agents and interactive characters. Like the original Chatterbox, it is MIT-licensed and English-focused.

Pi bon pou: Real-time voice agents, expressive speech with natural sounds

Navigue tout Chatterbox Turbo Voy

Yon ti gade

Pwogramè
Resemble AI
Lisans
MIT
Nivo
standard
Vitès
fast
Klonaj vwa
Wi
Lang
English
Karakteris maksimòm
1000

Chatterbox Turbo Voy

Default

English
Standart Neutral

Chatterbox Turbo TTS — FAQ

Turbo is a 350M-parameter model that runs at up to 6x real-time with sub-200ms latency, making it suitable for real-time voice agents where the original Chatterbox would be too slow.

They are inline cues like [laugh], [cough], and [chuckle] that you place directly in the text. The model performs the corresponding non-speech sounds, adding natural expressiveness.

Yes. All generated audio includes Perth watermarking, which supports provenance tracking and helps identify the output as AI-generated.
← Tout vwa