Chatterbox Turbo

Chatterbox Turbo TTS

A faster Chatterbox with sub-200ms latency and inline paralinguistic tags for laughs, coughs, and chuckles.

Regisztrálj! 5000 karakterhatárra

Írja be a szöveget az SSML címkékbe a pontos vezérlés érdekében:

<speak><prosody rate="slow">Slow speech</prosody></speak>

A kijelölt modell tagjeiben a ~ kattintson az alábbi szövegbe:

Ez a modell egyszerű szöveget olvas, így a sorcímke figyelmen kívül marad. A tag alapú érzelem, váltson át egy expresszív modell, mint Orpheus vagy Bark.

Definiáld az egyéni kiejtéseket (szó = kiejtés):

-12 +12
0.5x 2.0x
Szabad Piper, VITS, MelotTS
A generált audio jelenik meg itt. Válasszon ki egy modellt, írja be a szöveget, és kattintson a Generate gombra.
Audio generált sikeresen
0:00
Audio letöltése Letöltés.srt A kapcsolat 24 órán belül lejár
Ingyenes szint: személyes használat. Kereskedelmi engedély 5 dollárról
Mondd el a barátaidnak!

About Chatterbox Turbo

Chatterbox Turbo is Resemble AI's 350M-parameter speed-focused upgrade to Chatterbox, reaching up to 6x real-time generation with sub-200-millisecond latency. It keeps the original's voice cloning while adding inline paralinguistic tags — you can drop [laugh], [cough], or [chuckle] directly into your text and have the model perform them. Every generation carries Perth watermarking for provenance tracking, a nod to responsible-AI deployment. The combination of low latency and expressive non-speech sounds makes it well suited to real-time voice agents and interactive characters. Like the original Chatterbox, it is MIT-licensed and English-focused.

Legjobb: Real-time voice agents, expressive speech with natural sounds

Összes böngészés Chatterbox Turbo hangok

Egy pillantásra

Fejlesztő
Resemble AI
Jogosítvány
MIT
Tier
standard
Sebesség
fast
Hang klónozása
Igen.
Nyelvek
English
Max. karakterek
1000

Chatterbox Turbo hangok

Default

English
Szabvány Neutral

Chatterbox Turbo TTS - FAQCharselect unicode block name (optional, probably does not need a translation)

Turbo is a 350M-parameter model that runs at up to 6x real-time with sub-200ms latency, making it suitable for real-time voice agents where the original Chatterbox would be too slow.

They are inline cues like [laugh], [cough], and [chuckle] that you place directly in the text. The model performs the corresponding non-speech sounds, adding natural expressiveness.

Yes. All generated audio includes Perth watermarking, which supports provenance tracking and helps identify the output as AI-generated.
← Minden hang