Chatterbox Turbo

Chatterbox Turbo TTS

A faster Chatterbox with sub-200ms latency and inline paralinguistic tags for laughs, coughs, and chuckles.

Cofrestru am gyfyngiad 5,000 nod

Amlapio' ch testun mewn tagiau SSML er mwyn cael rheoli cywir:

<speak><prosody rate="slow">Slow speech</prosody></speak>

Tags y deall y model dewisiedig - cliciwch i daflu un i' ch testun lle mae' n digwydd:

Mae'r model yma yn darllen testun plaen, felly anwybyddir tagiau mewnlin. I ddelweddu teimlad yn seiliedig ar dagiau, newidiwch i ddelweddu mynegiant fel Orpheus neu Bark.

Diffinio ynganiad addasiedig (gair = ynganiad):

-12 +12
0.5x 2.0x
Am ddim gyda Piper, VITS, MeloTTS
Bydd eich sain a gynhyrchwyd yn ymddangos yma. Dewiswch ddull, rhowch destun, a chliciwch Creu.
Creuwyd Sain yn Llwyddiannus
0:00
Lawrlwytho Sain Lawrlwytho.srt Mae'r cyswllt yn darfod mewn 24 awr
Haen rhad: defnydd personol. Trwydded fasnachol o $5/mis
Gwneud hwn yn eich llais eich hun Clonio llais mewn 30 eiliad
Hoffwch TTS.ai? Meddwl am eich ffrindiau!

Am Chatterbox Turbo

Chatterbox Turbo is Resemble AI's 350M-parameter speed-focused upgrade to Chatterbox, reaching up to 6x real-time generation with sub-200-millisecond latency. It keeps the original's voice cloning while adding inline paralinguistic tags — you can drop [laugh], [cough], or [chuckle] directly into your text and have the model perform them. Every generation carries Perth watermarking for provenance tracking, a nod to responsible-AI deployment. The combination of low latency and expressive non-speech sounds makes it well suited to real-time voice agents and interactive characters. Like the original Chatterbox, it is MIT-licensed and English-focused.

Gorau ar gyfer: Real-time voice agents, expressive speech with natural sounds

Pori Popeth Chatterbox Turbo Saesneg

Yn syth

Datblygwr
Resemble AI
Trwydded
MIT
o Fawrth
standard
Cyflymder
fast
Clonio llais
IeQShortcut
Iaith:
English
Uchafswm nodau
1000

Chatterbox Turbo Saesneg

Default

English
Arferol Neutral

Chatterbox Turbo TTS - Cwestiynau Cyffredin

Turbo is a 350M-parameter model that runs at up to 6x real-time with sub-200ms latency, making it suitable for real-time voice agents where the original Chatterbox would be too slow.

They are inline cues like [laugh], [cough], and [chuckle] that you place directly in the text. The model performs the corresponding non-speech sounds, adding natural expressiveness.

Yes. All generated audio includes Perth watermarking, which supports provenance tracking and helps identify the output as AI-generated.
← Pob llais