Chatterbox Turbo

Chatterbox Turbo TTS

A faster Chatterbox with sub-200ms latency and inline paralinguistic tags for laughs, coughs, and chuckles.

Zarejestruj się. dla 5000 limitów znaków

Zawiń tekst w tagi SSML dla precyzyjnej kontroli:

<speak><prosody rate="slow">Slow speech</prosody></speak>

Tagi wybrany model rozumie — kliknij aby usunąć jeden do swojego tekstu, gdzie się to dzieje:

Model ten czyta tekst zwykły, więc w linii tagi są ignorowane. Dla emocji na tag, przełącz na model wyrażony jak Orfeus lub Bark.

Definiuj własny wymówki (słowo = wymówka):

-12 +12
0.5x 2.0x
Darmowe z Piper, VITS, Melotts
Tutaj pojawi się generowany dźwięk. Wybierz model, wpisz tekst i kliknij Generuj.
Pomyślnie wygenerowany dźwięk
0:00
Pobierz audio Pobierz.rt Łączność wygasa w 24h
Bezpłatny poziom: użytkowanie osobiste. Licencja handlowa od $5/mo
Powiedz znajomym!

O tematie Chatterbox Turbo

Chatterbox Turbo is Resemble AI's 350M-parameter speed-focused upgrade to Chatterbox, reaching up to 6x real-time generation with sub-200-millisecond latency. It keeps the original's voice cloning while adding inline paralinguistic tags — you can drop [laugh], [cough], or [chuckle] directly into your text and have the model perform them. Every generation carries Perth watermarking for provenance tracking, a nod to responsible-AI deployment. The combination of low latency and expressive non-speech sounds makes it well suited to real-time voice agents and interactive characters. Like the original Chatterbox, it is MIT-licensed and English-focused.

Najlepsze dla: Real-time voice agents, expressive speech with natural sounds

Przeglądaj wszystkie Chatterbox Turbo głosy

Na jedno spojrzenie

Rozwijacz
Resemble AI
Licencja
MIT
Poziom szczelności
standard
Prędkość
fast
Klonowanie głosu
Tak.
Języki
English
Maksymalna liczba znaków
1000

Chatterbox Turbo głosy

Default

English
Standardowe Neutral

Chatterbox Turbo TTS — FAQ

Turbo is a 350M-parameter model that runs at up to 6x real-time with sub-200ms latency, making it suitable for real-time voice agents where the original Chatterbox would be too slow.

They are inline cues like [laugh], [cough], and [chuckle] that you place directly in the text. The model performs the corresponding non-speech sounds, adding natural expressiveness.

Yes. All generated audio includes Perth watermarking, which supports provenance tracking and helps identify the output as AI-generated.
← Wszystkie głosy