Chatterbox Turbo

Chatterbox Turbo TTS

A faster Chatterbox with sub-200ms latency and inline paralinguistic tags for laughs, coughs, and chuckles.

Prihlásiť sa na odber Limit 5 000 znakov

Zabaliť text do SSML značiek pre presnú kontrolu:

<speak><prosody rate="slow">Slow speech</prosody></speak>

Značky, ktorým vybraný model rozumie — kliknutím ich umiestnite do textu tam, kde sa vyskytujú:

Tento model číta obyčajný text, takže vnorené značky sa ignorujú.Pre emócie založené na značkách prejdite na expresívny model ako Orpheus alebo Bark.

Definovať vlastné výslovnosti (slovo = výslovnosť):

-12 +12
0.5x 2.0x
Zadarmo s Piper, VITS, MeloTTS
Vyberte si model, zadajte text a kliknite na tlačidlo Generovať.Generate.
Audio generované úspešne
0:00
Stiahnuť audio na stiahnutie Stiahnuť.srt súbor Platnosť odkazu vyprší za 24h
Bezplatná úroveň: osobné použitie. Komerčná licencia od 5 USD/mesiac
Láska TTS.ai? Povedzte svojim priateľom!

O nás Chatterbox Turbo

Chatterbox Turbo is Resemble AI's 350M-parameter speed-focused upgrade to Chatterbox, reaching up to 6x real-time generation with sub-200-millisecond latency. It keeps the original's voice cloning while adding inline paralinguistic tags — you can drop [laugh], [cough], or [chuckle] directly into your text and have the model perform them. Every generation carries Perth watermarking for provenance tracking, a nod to responsible-AI deployment. The combination of low latency and expressive non-speech sounds makes it well suited to real-time voice agents and interactive characters. Like the original Chatterbox, it is MIT-licensed and English-focused.

Najlepšie pre: Real-time voice agents, expressive speech with natural sounds

Prehľadávať všetky Chatterbox Turbo hlasy

Na prvý pohľad

Vývojár
Resemble AI
Licencia
MIT
Zvieratá
standard
Rýchlosť
fast
Klonovanie hlasu
Áno
Jazyky
English
Max. počet znakov
1000

Chatterbox Turbo hlasy

Default

English
Štandardné Neutral

Chatterbox Turbo TTS — Často kladené otázky

Turbo is a 350M-parameter model that runs at up to 6x real-time with sub-200ms latency, making it suitable for real-time voice agents where the original Chatterbox would be too slow.

They are inline cues like [laugh], [cough], and [chuckle] that you place directly in the text. The model performs the corresponding non-speech sounds, adding natural expressiveness.

Yes. All generated audio includes Perth watermarking, which supports provenance tracking and helps identify the output as AI-generated.
← Všetky hlasy