Darwin TTS

Darwin TTS TTS

A Qwen3-TTS variant whose talker FFN weights are blended from the Qwen3 language model for sharper cross-lingual cloning.

Pierakstīties 5000 rakstzīmju limitam

Aplauzt savu tekstu SSML tagus precīzai kontrolei:

<speak><prosody rate="slow">Slow speech</prosody></speak>

Tags izvēlētais modelis saprot — noklikšķiniet, lai iemestu vienu jūsu tekstā, kur tas notiek:

Šis modelis lasa vienkāršu tekstu, tāpēc tiek ignorēti inline tagi. Uz tag-based emocijas, pāriet uz izteiksmīgu modeli, piemēram, Orpheus vai Bark.

Definēt pielāgotu izrunas (vārds = izruna):

-12 +12
0.5x 2.0x
Bez piper, VITS, MeloTTS
Šeit parādīsies jūsu ģenerētais audio. Izvēlieties modeli, ievadiet tekstu un noklikšķiniet ģenerējiet.
Audio veiksmīgi ģenerēts
0:00
Lejupielādēt audio Lejupielādēt.srt Saite beidzas 24h
Bezmaksas līmenis: personīgai lietošanai. Komerclicenci no $5/mo
Mīlestība TTS.ai? Stāsti saviem draugiem!

Par Darwin TTS

Darwin-TTS-1.7B-Cross by FINAL-Bench is a research variant of Qwen3-TTS-1.7B with an unusual construction: 84 of its talker-FFN tensors (about 8.6% of them) are blended at a 3% ratio with the matching tensors from Qwen3-1.7B-Base, all without any retraining. The result is a model that produces noticeably crisper cross-lingual voice cloning across Korean, English, Japanese, and Chinese — its four core languages. It operates in zero-shot voice-clone mode, needing only about three seconds of reference audio to capture a speaker. Darwin is best suited to transferring a single reference voice across those four languages, for example dubbing or multilingual narration with consistent speaker identity.

Labākais: Cross-lingual voice cloning between English / Korean / Japanese / Chinese with a single reference voice

Pārlūkot visu Darwin TTS balsis

Īsumā

Izstrādātājs
FINAL-Bench
Licence
Apache 2.0
Līmeņrādis
standard
Ātrums
medium
Balss klonēšana
Valodas
English, Korean, Japanese, Chinese
Maks. rakstzīmes
2000

Darwin TTS balsis

Default

English
Standarta Neutral

Default (Chinese)

Chinese
Standarta Neutral

Default (Japanese)

Japanese
Standarta Neutral

Default (Korean)

Korean
Standarta Neutral

Darwin TTS TTS – FAQ

Darwin starts from Qwen3-TTS-1.7B but blends a small fraction of its talker-FFN weights with the matching weights from the Qwen3-1.7B base language model. This training-free blend sharpens cross-lingual voice cloning rather than changing the base voices.

English, Korean, Japanese, and Chinese. The FINAL-Bench release specifically markets its cross-lingual blend for those four, and the deployed model ships voices for them.

About three seconds. It works in zero-shot mode, so no fine-tuning or training is required — you provide a short reference clip and it generates new speech in that voice.
← Visas balsis