Chatterbox

Chatterbox TTS

Resemble AI's state-of-the-art zero-shot voice cloning model with independent emotion control.

Melden Sie sich an für 5.000 Zeichen-Grenze

Verpacken Sie Ihren Text in SSML-Tags für eine präzise Kontrolle:

<speak><prosody rate="slow">Slow speech</prosody></speak>

Tags, die das ausgewählte Modell versteht — klicken Sie, um einen in Ihren Text zu legen, wo es passiert:

Dieses Modell liest Text, so dass Inline-Tags ignoriert werden. Für tag-basierte Emotion, wechseln Sie zu einem ausdrucksstarken Modell wie Orpheus oder Bark.

Benutzerdefinierte Aussprachen definieren (Wort = Aussprache):

-12 +12
0.5x 2.0x
Frei mit Piper, VITS, MeloTTS
Hier erscheint Ihr generiertes Audio. Wählen Sie ein Modell, geben Sie Text ein und klicken Sie auf Generieren.
Audio-Erzeugung erfolgreich
0:00
Audio herunterladen Download.srt Link läuft in 24h aus
Freier Dienstgrad: persönlicher Gebrauch. Kommerzielle Lizenz ab $5/mo
Gefällt dir TTS.ai? Erzähl es deinen Freunden!

Über Chatterbox

Chatterbox by Resemble AI is a leading open-source zero-shot voice cloning model that replicates a voice from a single audio sample, capturing not just timbre but speaking style and emotional nuance. Its distinctive feature is fine-grained emotion control that operates independently of the voice identity, so you can keep a cloned voice but shift its emotional intensity. Built around ResembleEnhance and flow matching, it targets professional-grade cloning for content creation, dubbing, and character voices. Released under the permissive MIT license, Chatterbox has become a popular foundation for derivative models — TTS.ai also runs a Saudi-Arabic fine-tune of it. It favors quality, with a modest per-request character limit.

Das Beste für: Professional voice cloning with emotional control, content creation

Alle durchsuchen Chatterbox Stimmen

Auf einen Blick

Entwickler
Resemble AI
Lizenz
MIT
Tierart
premium
Geschwindigkeit
medium
Klonen der Stimme
Nein
Sprachen
English
Maximale Zeichen
300

Chatterbox Stimmen

Default

English
Prämie Neutral

Chatterbox TTS — FAQ

A single audio sample is enough. Chatterbox is a zero-shot cloning model, so it captures a voice — including its style and emotional nuance — from one reference clip without any fine-tuning.

It offers fine-grained emotion control that works independently from the voice identity, letting you adjust the emotional tone of the output while keeping the same cloned voice.

Yes. Chatterbox is released by Resemble AI under the MIT license, which permits commercial use, and it serves as the base for several fine-tuned derivative models.
← Alle Stimmen