Chatterbox Turbo

Chatterbox Turbo TTS

A faster Chatterbox with sub-200ms latency and inline paralinguistic tags for laughs, coughs, and chuckles.

Գրանցվել 5000 սանտիմետր սահմանափակում

Ձեր տեքստը SSML տեգերի մեջ տեղադրել ճշգրիտ կառավարման համար.

<speak><prosody rate="slow">Slow speech</prosody></speak>

Ընտրված մոդելի կողմից ընկալվող պիտակներ — սեղմեք դրանք ձեր տեքստում տեղադրելու համար:

Այս մոդելը կարդում է պարզ տեքստը, այնպես որ ներառված տեքստերը անտեսվում են։ Տեքստերի վրա հիմնված էմոցիաների համար փոխեք արտահայտիչ մոդել, ինչպես Orpheus կամ Bark։

Որոշել սեփական արտասանությունը (բառ = արտասանություն):

-12 +12
0.5x 2.0x
Ազատ Piper, VITS, MeloTTS-ով
Այստեղ կհայտնվի ձեր ստեղծած ձայնը։ Ընտրեք մոդել, ներդրեք տեքստ և սեղմեք Ծնվել։
Ավտոմատ ձայնագրում
0:00
Տեղադրել ձայնային Տեղադրել.srt Հղումն ավարտվում է 24 ժամ անց
Ֆրանսիա: Ֆրանսիայի ազգային ռադիո: Ֆրանսիայի ազգային ռադիո. Կազմակերպական լիցենզիա $5/ամս
Սիրում եք TTS.ai-ն? Պատմեք ձեր ընկերներին։

Ընդհանուր Chatterbox Turbo

Chatterbox Turbo is Resemble AI's 350M-parameter speed-focused upgrade to Chatterbox, reaching up to 6x real-time generation with sub-200-millisecond latency. It keeps the original's voice cloning while adding inline paralinguistic tags — you can drop [laugh], [cough], or [chuckle] directly into your text and have the model perform them. Every generation carries Perth watermarking for provenance tracking, a nod to responsible-AI deployment. The combination of low latency and expressive non-speech sounds makes it well suited to real-time voice agents and interactive characters. Like the original Chatterbox, it is MIT-licensed and English-focused.

Լավագույնը: Real-time voice agents, expressive speech with natural sounds

Ընթերցել բոլորը Chatterbox Turbo ձայներ

Հակառակորդի դիրքը

Հեղինակ
Resemble AI
Լիցենզիա
MIT
Դադար
standard
արագություն
fast
Ձայնի կլոնավորում
Այո
Լեզուներ
English
Օգտագործված ռեժիմ
1000

Chatterbox Turbo ձայներ

Default

English
Լռելյայն Neutral

Chatterbox Turbo TTS - Հաճախ տրվող հարցեր

Turbo is a 350M-parameter model that runs at up to 6x real-time with sub-200ms latency, making it suitable for real-time voice agents where the original Chatterbox would be too slow.

They are inline cues like [laugh], [cough], and [chuckle] that you place directly in the text. The model performs the corresponding non-speech sounds, adding natural expressiveness.

Yes. All generated audio includes Perth watermarking, which supports provenance tracking and helps identify the output as AI-generated.
← Բոլոր ձայները