Chatterbox Turbo

Chatterbox Turbo TTS

A faster Chatterbox with sub-200ms latency and inline paralinguistic tags for laughs, coughs, and chuckles.

Zaregistrovat se pro 5000 znaků limit

Zabalte svůj text do značek SSML pro přesné ovládání:

<speak><prosody rate="slow">Slow speech</prosody></speak>

Značky vybraného modelu rozumí? klikněte na tlačítko pro kapku jednoho do textu, kde se to stane:

Tento model čte prostý text, takže inline značky jsou ignorovány. Pro tag-based emotion, přepněte na expresivní model, jako je Orpheus nebo Bark.

Definovat vlastní výslovnosti (slovo = výslovnost):

-12 +12
0.5x 2.0x
Zdarma s Piper, VITS, MeloTTS
Zde se objeví váš vygenerovaný zvuk. Vyberte model, zadejte text a klikněte na Generovat.
Audio generované úspěšně
0:00
Stáhnout zvuk Stáhnout.srt Odkaz vyprší v 24 hodin
Volný stupeň: osobní použití. Obchodní licence od 5 dolarů/mo
Miluju TTS.ai? Řekni to svým přátelům!

O aplikaci Chatterbox Turbo

Chatterbox Turbo is Resemble AI's 350M-parameter speed-focused upgrade to Chatterbox, reaching up to 6x real-time generation with sub-200-millisecond latency. It keeps the original's voice cloning while adding inline paralinguistic tags — you can drop [laugh], [cough], or [chuckle] directly into your text and have the model perform them. Every generation carries Perth watermarking for provenance tracking, a nod to responsible-AI deployment. The combination of low latency and expressive non-speech sounds makes it well suited to real-time voice agents and interactive characters. Like the original Chatterbox, it is MIT-licensed and English-focused.

Nejlepší pro: Real-time voice agents, expressive speech with natural sounds

Procházet vše Chatterbox Turbo hlasy

Na první pohled

Vývojář
Resemble AI
Licence
MIT
Úroveň
standard
Rychlost
fast
Klonování hlasu
Ano.
Jazyky
English
Max znaků
1000

Chatterbox Turbo hlasy

Default

English
Standardní Neutral

Chatterbox Turbo FAQ TTS

Turbo is a 350M-parameter model that runs at up to 6x real-time with sub-200ms latency, making it suitable for real-time voice agents where the original Chatterbox would be too slow.

They are inline cues like [laugh], [cough], and [chuckle] that you place directly in the text. The model performs the corresponding non-speech sounds, adding natural expressiveness.

Yes. All generated audio includes Perth watermarking, which supports provenance tracking and helps identify the output as AI-generated.
← Všechny hlasy