Chatterbox Turbo

Chatterbox Turbo TTS

A faster Chatterbox with sub-200ms latency and inline paralinguistic tags for laughs, coughs, and chuckles.

Daftar untuk batas 5,000 karakter

Bungkus teks Anda dalam tag SSML untuk kendali yang tepat:

<speak><prosody rate="slow">Slow speech</prosody></speak>

Tag yang dipilih mengerti klik °C untuk memasukkan satu ke dalam teks Anda di mana hal itu terjadi:

Model ini membaca teks biasa, sehingga tag inline diabaikan. Untuk tag berbasis emosi, beralih ke model ekspresif seperti Orpheus atau Bark.

Definisikan pengucapan ubahan (kata = pelafalan):

-12 +12
0.5x 2.0x
Free with Piper, VITS, Melotts
Audio yang Anda buat akan muncul di sini. Pilih model, masukkan teks, dan klik Generate.
Hasil Audio Berhasil
0:00
Unduh Audio Unduh.srt Sambungan berakhir dalam 24 jam
Tingkatan bebas: penggunaan pribadi. Ijin komersial dari $5/mo
Buatlah ini suara Anda sendiri Kloning suara dalam 30 detik
Beritahu teman-temanmu!

Tentang Chatterbox Turbo

Chatterbox Turbo is Resemble AI's 350M-parameter speed-focused upgrade to Chatterbox, reaching up to 6x real-time generation with sub-200-millisecond latency. It keeps the original's voice cloning while adding inline paralinguistic tags — you can drop [laugh], [cough], or [chuckle] directly into your text and have the model perform them. Every generation carries Perth watermarking for provenance tracking, a nod to responsible-AI deployment. The combination of low latency and expressive non-speech sounds makes it well suited to real-time voice agents and interactive characters. Like the original Chatterbox, it is MIT-licensed and English-focused.

Terbaik untuk: Real-time voice agents, expressive speech with natural sounds

Jelajahi semua Chatterbox Turbo suara

Pada sekilas

Pengembang
Resemble AI
Lisensi
MIT
Tier
standard
Kecepatan
fast
Penklonan Suara
Ya
Bahasa
English
Karakter maksimal
1000

Chatterbox Turbo suara

Default

English
Standar Neutral

Chatterbox Turbo TTS °F FAQ

Turbo is a 350M-parameter model that runs at up to 6x real-time with sub-200ms latency, making it suitable for real-time voice agents where the original Chatterbox would be too slow.

They are inline cues like [laugh], [cough], and [chuckle] that you place directly in the text. The model performs the corresponding non-speech sounds, adding natural expressiveness.

Yes. All generated audio includes Perth watermarking, which supports provenance tracking and helps identify the output as AI-generated.
← Semua suara