Chatterbox Turbo

Chatterbox Turbo TTS

A faster Chatterbox with sub-200ms latency and inline paralinguistic tags for laughs, coughs, and chuckles.

ਸਾਈਨ ਅੱਪ 5, 000 ਅੱਖਰ ਲਿਮਟ

ਸਹੀ ਕੰਟਰੋਲ ਲਈ SSML ਟੈਗ ਵਿੱਚ ਆਪਣਾ ਪਾਠ ਲਪੇਟੋ:

<speak><prosody rate="slow">Slow speech</prosody></speak>

ਟੈਗ, ਜੋ ਕਿ ਚੁਣੇ ਮਾਡਲ ਸਮਝਦਾ ਹੈ - ਆਪਣੇ ਪਾਠ ਵਿੱਚ ਇੱਕ ਟੈਗ ਸੁੱਟਣ ਲਈ ਕਲਿੱਕ ਕਰੋ, ਜਿੱਥੇ ਇਹ ਹੁੰਦਾ ਹੈ:

ਇਹ ਮਾਡਲ ਸਾਦਾ ਪਾਠ ਪੜ੍ਹਦਾ ਹੈ, ਇਸ ਲਈ ਇੰਲਾਈਨ ਟੈਗ ਅਣਡਿੱਠੇ ਕੀਤੇ ਜਾਂਦੇ ਹਨ । ਟੈਗ- ਅਧਾਰਿਤ ਈਮੋਸ਼ਨ ਲਈ ਇੱਕ ਸਪੱਸ਼ਟ ਮਾਡਲ ਜਿਵੇਂ ਕਿ Orpheus ਜਾਂ Bark ਲਈ ਬਦਲੋ ।

ਪਸੰਦੀਦਾ ਉਚਾਰਨ ਦਿਓ (ਸ਼ਬਦ = ਉਚਾਰਨ):

-12 +12
0.5x 2.0x
ਪਾਈਪਰ, VITS, MeloTTS ਨਾਲ ਮੁਫਤ
ਤੁਹਾਡਾ ਬਣਾਇਆ ਆਡੀਓ ਇੱਥੇ ਵੇਖਾਇਆ ਜਾਵੇਗਾ । ਇੱਕ ਮਾਡਲ ਚੁਣੋ, ਪਾਠ ਦਿਓ ਅਤੇ ਬਣਾਓ ਕਲਿੱਕ ਕਰੋ ।
ਆਡੀਓ ਸਫਲਤਾਪੂਰਕ ਬਣਾਇਆ ਗਿਆ
0:00
ਆਡੀਓ ਡਾਊਨਲੋਡ .srt ਡਾਊਨਲੋਡ ਲਿੰਕ 24 ਘੰਟਿਆਂ ਵਿੱਚ ਖਤਮ ਹੁੰਦਾ ਹੈ
ਮੁਫਤ ਪੱਧਰ: ਨਿੱਜੀ ਵਰਤੋਂ ਲਈ। $5/ਮਹੀਨੇ ਤੋਂ ਵਪਾਰਕ ਲਾਇਸੈਂਸ
TTS.ai ਪਸੰਦ ਹੈ? ਆਪਣੇ ਦੋਸਤਾਂ ਨੂੰ ਦੱਸੋ!

ਬਾਰੇ Chatterbox Turbo

Chatterbox Turbo is Resemble AI's 350M-parameter speed-focused upgrade to Chatterbox, reaching up to 6x real-time generation with sub-200-millisecond latency. It keeps the original's voice cloning while adding inline paralinguistic tags — you can drop [laugh], [cough], or [chuckle] directly into your text and have the model perform them. Every generation carries Perth watermarking for provenance tracking, a nod to responsible-AI deployment. The combination of low latency and expressive non-speech sounds makes it well suited to real-time voice agents and interactive characters. Like the original Chatterbox, it is MIT-licensed and English-focused.

ਇਸ ਲਈ ਸਭ ਤੋਂ ਵਧੀਆ: Real-time voice agents, expressive speech with natural sounds

ਸਭ ਝਲਕ Chatterbox Turbo ਆਵਾਜ਼ਾਂ

ਇੱਕ ਨਜ਼ਰ

ਡਿਵੈਲਪਰ
Resemble AI
ਲਾਈਸੈਂਸ
MIT
ਟੀਅਰ
standard
ਗਤੀ
fast
ਬੋਲੀ ਕਲੋਨਿੰਗ
ਹਾਂ
ਭਾਸ਼ਾਵਾਂ
English
ਵੱਧੋ- ਵੱਧ ਅੱਖਰ
1000

Chatterbox Turbo ਆਵਾਜ਼ਾਂ

Default

English
ਸਟੈਂਡਰਡ Neutral

Chatterbox Turbo TTS - ਅਕਸਰ ਪੁੱਛੇ ਜਾਂਦੇ ਸਵਾਲ

Turbo is a 350M-parameter model that runs at up to 6x real-time with sub-200ms latency, making it suitable for real-time voice agents where the original Chatterbox would be too slow.

They are inline cues like [laugh], [cough], and [chuckle] that you place directly in the text. The model performs the corresponding non-speech sounds, adding natural expressiveness.

Yes. All generated audio includes Perth watermarking, which supports provenance tracking and helps identify the output as AI-generated.
← ਸਭ ਆਵਾਜ਼ਾਂ