Chatterbox

Chatterbox TTS

Resemble AI's state-of-the-art zero-shot voice cloning model with independent emotion control.

Kugadzwa for 5,000 character limit

Wrap yako tenzi mu SSML tags kuti zvive nyore kudzora:

<speak><prosody rate="slow">Slow speech</prosody></speak>

Tags iyo yakasarudzwa model inonzwisiswa — tinya kuti utore imwe muchinyorwa chako apo inoitika:

Iyi modhi inoverenga chete mazita ezvinyorwa, saka zvinyorwa zvinonyorwa mumitsara hazvina kutariswa. Kuti uwane pfungwa dziri mumitsara, chinja kune imwe modhi inoratidza pfungwa seOrpheus kana Bark.

Define custom pronunciations (word = pronunciation):

-12 +12
0.5x 2.0x
Free with Piper, VITS, MeloTTS
Yako yakagadzirwa audio ichaonekwa pano. Choose a model, enter text, and click Generate.
Audio Yakagadzirwa Nekubudirira
0:00
Download Audio Dhawunirodha.srt Link inotanga kushanda mu 24h
Free tier: personal usage. Commercial license from $5/mo
Love TTS.ai? Tiudza shamwari dzako!

Chii Chatterbox

Chatterbox by Resemble AI is a leading open-source zero-shot voice cloning model that replicates a voice from a single audio sample, capturing not just timbre but speaking style and emotional nuance. Its distinctive feature is fine-grained emotion control that operates independently of the voice identity, so you can keep a cloned voice but shift its emotional intensity. Built around ResembleEnhance and flow matching, it targets professional-grade cloning for content creation, dubbing, and character voices. Released under the permissive MIT license, Chatterbox has become a popular foundation for derivative models — TTS.ai also runs a Saudi-Arabic fine-tune of it. It favors quality, with a modest per-request character limit.

Yakanaka kune: Professional voice cloning with emotional control, content creation

Tarisa zvese Chatterbox mazwi

Mufananidzo

Developer
Resemble AI
License
MIT
Tier
premium
Speed
medium
Kutaura
Yes
Zvinhu
English
Max characters
300

Chatterbox mazwi

Default

English
Premium Neutral

Chatterbox TTS — Zvinyorwa

A single audio sample is enough. Chatterbox is a zero-shot cloning model, so it captures a voice — including its style and emotional nuance — from one reference clip without any fine-tuning.

It offers fine-grained emotion control that works independently from the voice identity, letting you adjust the emotional tone of the output while keeping the same cloned voice.

Yes. Chatterbox is released by Resemble AI under the MIT license, which permits commercial use, and it serves as the base for several fine-tuned derivative models.
← All voices