Chatterbox Turbo

Chatterbox Turbo TTS

A faster Chatterbox with sub-200ms latency and inline paralinguistic tags for laughs, coughs, and chuckles.

Aliĝi for 5, 000 character limit

Envolvu vian tekston en SSML- etikedojn por preciza kontrolo:

<speak><prosody rate="slow">Slow speech</prosody></speak>

Etikedoj kiujn la elektita modelo komprenas - klaku por meti unu en vian tekston kie ĝi okazas:

This model reads plain text, so inline tags are ignored. For tag-based emotion, switch to an expressive model like Orpheus or Bark.

Difini proprajn elparolojn (vorto = elparolo):

-12 +12
0.5x 2.0x
Libera kun Piper, VITS, MeloTTS
Via generita sono aperos tie ĉi. Elektu modelon, entajpu tekston, kaj alklaku Generi.
Sondosiero sukcese generita
0:00
Elŝuti sonon Elŝuti.srt Ligo eksvalidiĝas post 24 horoj
Libera programaro: persona uzo. Komerca licenco ekde $5/mo
Ĉu vi ŝatas TTS.ai? Diru al viaj amikoj!

Pri Chatterbox Turbo

Chatterbox Turbo is Resemble AI's 350M-parameter speed-focused upgrade to Chatterbox, reaching up to 6x real-time generation with sub-200-millisecond latency. It keeps the original's voice cloning while adding inline paralinguistic tags — you can drop [laugh], [cough], or [chuckle] directly into your text and have the model perform them. Every generation carries Perth watermarking for provenance tracking, a nod to responsible-AI deployment. The combination of low latency and expressive non-speech sounds makes it well suited to real-time voice agents and interactive characters. Like the original Chatterbox, it is MIT-licensed and English-focused.

Plej bona por: Real-time voice agents, expressive speech with natural sounds

Foliumi ĉiujn Chatterbox Turbo voĉoj

Unu rigardo

Programisto
Resemble AI
Licenco
MIT
Tamuz
standard
Rapideco
fast
Voĉo- klonado
Jes
Lingvoj
English
Maksimuma nombro da signoj
1000

Chatterbox Turbo voĉoj

Default

English
Defaŭlta Neutral

Chatterbox Turbo TTS - FAQ

Turbo is a 350M-parameter model that runs at up to 6x real-time with sub-200ms latency, making it suitable for real-time voice agents where the original Chatterbox would be too slow.

They are inline cues like [laugh], [cough], and [chuckle] that you place directly in the text. The model performs the corresponding non-speech sounds, adding natural expressiveness.

Yes. All generated audio includes Perth watermarking, which supports provenance tracking and helps identify the output as AI-generated.
← Ĉiuj voĉoj