Chatterbox Turbo

Chatterbox Turbo TTS

A faster Chatterbox with sub-200ms latency and inline paralinguistic tags for laughs, coughs, and chuckles.

Iscriviti per un limite di 5.000 caratteri

Avvolgi il tuo testo nei tag SSML per un controllo preciso:

<speak><prosody rate="slow">Slow speech</prosody></speak>

Tags il modello selezionato comprende clic su

Questo modello legge testo semplice, quindi i tag inline vengono ignorati. Per le emozioni basate sui tag, passare a un modello espressivo come Orpheus o Bark.

Definire le pronunciazioni personalizzate (parola = pronuncia):

-12 +12
0.5x 2.0x
Gratis con Piper, VITS, MeloTTS
L'audio generato apparirà qui. Scegli un modello, inserisci testo e fai clic su Genera.
Audio generato con successo
0:00
Scarica audio Scarica.srt Link scade in 24 ore
Livello libero: uso personale. Licenza commerciale da $5/mo
Fai di questo la tua voce Clona una voce in 30 secondi
Ti piace TTS.ai? Dillo ai tuoi amici!

Informazioni Chatterbox Turbo

Chatterbox Turbo is Resemble AI's 350M-parameter speed-focused upgrade to Chatterbox, reaching up to 6x real-time generation with sub-200-millisecond latency. It keeps the original's voice cloning while adding inline paralinguistic tags — you can drop [laugh], [cough], or [chuckle] directly into your text and have the model perform them. Every generation carries Perth watermarking for provenance tracking, a nod to responsible-AI deployment. The combination of low latency and expressive non-speech sounds makes it well suited to real-time voice agents and interactive characters. Like the original Chatterbox, it is MIT-licensed and English-focused.

Meglio per: Real-time voice agents, expressive speech with natural sounds

Sfoglia tutti Chatterbox Turbo voci

A colpo d'occhio

Sviluppatore
Resemble AI
Licenza
MIT
Livello
standard
Velocità
fast
Clonazione vocale
Lingue
English
Caratteri massimi
1000

Chatterbox Turbo voci

Default

English
Standard Neutral

Chatterbox Turbo FAQ del TTS

Turbo is a 350M-parameter model that runs at up to 6x real-time with sub-200ms latency, making it suitable for real-time voice agents where the original Chatterbox would be too slow.

They are inline cues like [laugh], [cough], and [chuckle] that you place directly in the text. The model performs the corresponding non-speech sounds, adding natural expressiveness.

Yes. All generated audio includes Perth watermarking, which supports provenance tracking and helps identify the output as AI-generated.
← Tutte le voci