Chatterbox Turbo

Chatterbox Turbo TTS

A faster Chatterbox with sub-200ms latency and inline paralinguistic tags for laughs, coughs, and chuckles.

Inscríbete para el límite de 5.000 caracteres

Envuelva su texto en etiquetas SSML para un control preciso:

<speak><prosody rate="slow">Slow speech</prosody></speak>

Etiquetas el modelo seleccionado entiende — haga clic para soltar uno en su texto donde sucede:

Este modelo lee texto plano, por lo que las etiquetas en línea son ignoradas. Para la emoción basada en etiquetas, cambie a un modelo expresivo como Orfeo o Bark.

Definir pronunciaciones personalizadas (palabra = pronunciación):

-12 +12
0.5x 2.0x
Libre con Piper, VITS, MeloTTS
Su audio generado aparecerá aquí. Elija un modelo, introduzca texto y haga clic en Generar.
Audio generado con éxito
0:00
Descargar audio Descargar.srt Enlace expira en 24h
Nivel libre: uso personal. Licencia comercial desde $5/mes
Haz de esto tu propia voz Clonar una voz en 30 segundos
¿Te gusta TTS.ai? ¡Cuéntaselo a tus amigos!

Acerca de Chatterbox Turbo

Chatterbox Turbo is Resemble AI's 350M-parameter speed-focused upgrade to Chatterbox, reaching up to 6x real-time generation with sub-200-millisecond latency. It keeps the original's voice cloning while adding inline paralinguistic tags — you can drop [laugh], [cough], or [chuckle] directly into your text and have the model perform them. Every generation carries Perth watermarking for provenance tracking, a nod to responsible-AI deployment. The combination of low latency and expressive non-speech sounds makes it well suited to real-time voice agents and interactive characters. Like the original Chatterbox, it is MIT-licensed and English-focused.

Lo mejor para: Real-time voice agents, expressive speech with natural sounds

Examinar todo Chatterbox Turbo voces

De un vistazo

Desarrollador
Resemble AI
Licencia
MIT
Nivel
standard
Velocidad
fast
Clonación de voz
Idiomas
English
Máx. caracteres
1000

Chatterbox Turbo voces

Default

English
Estándar Neutral

Chatterbox Turbo TTS — Preguntas más frecuentes

Turbo is a 350M-parameter model that runs at up to 6x real-time with sub-200ms latency, making it suitable for real-time voice agents where the original Chatterbox would be too slow.

They are inline cues like [laugh], [cough], and [chuckle] that you place directly in the text. The model performs the corresponding non-speech sounds, adding natural expressiveness.

Yes. All generated audio includes Perth watermarking, which supports provenance tracking and helps identify the output as AI-generated.
← Todas las voces