Darwin TTS

Darwin TTS TTS

A Qwen3-TTS variant whose talker FFN weights are blended from the Qwen3 language model for sharper cross-lingual cloning.

Ku soo biir 5,000 xaraf xaddid

Wrap qoraalka ku SSML tags si loo hubiyo xakamaynta saxda ah:

<speak><prosody rate="slow">Slow speech</prosody></speak>

Tags qaabka la doortay fahmo — riix si aad u hoos mid ka mid ah qoraalka aad halkaas oo uu ka dhacaa:

Model this akhriyo qoraalka caadiga ah, sidaas inline tags waa la iska indho tiri. For tag-ku salaysan dareenka, u dhaqaaqo si ay u muujiyaan qaabka sida Orpheus ama Bark.

Define custom pronunciations (word = dhawaaqa):

-12 +12
0.5x 2.0x
Bilaash ah oo leh Piper, VITS, MeloTTS
Your audio soo saaro halkan ka muuqan doonaa. Dooro qaab, ku qor qoraalka, oo guji soo saaro.
Dhaqdhaqaaqa ayaa la soo saaray
0:00
Soo dejisa Soo dejisan.srt Xidhiidhku wuxuu dhamaaday 24h
Free tier: isticmaalka shakhsiga ah. Liisan ganacsi laga bilaabo $ 5 / mo
Jecel TTS.ai? Ka warran saaxiibadaa!

About Darwin TTS

Darwin-TTS-1.7B-Cross by FINAL-Bench is a research variant of Qwen3-TTS-1.7B with an unusual construction: 84 of its talker-FFN tensors (about 8.6% of them) are blended at a 3% ratio with the matching tensors from Qwen3-1.7B-Base, all without any retraining. The result is a model that produces noticeably crisper cross-lingual voice cloning across Korean, English, Japanese, and Chinese — its four core languages. It operates in zero-shot voice-clone mode, needing only about three seconds of reference audio to capture a speaker. Darwin is best suited to transferring a single reference voice across those four languages, for example dubbing or multilingual narration with consistent speaker identity.

Ugu Fiican: Cross-lingual voice cloning between English / Korean / Japanese / Chinese with a single reference voice

Taabo oo kala soo bax Darwin TTS cod

Eeg

Soo-saarayaasha
FINAL-Bench
Liisan
Apache 2.0
Qiyaamaha
standard
Xawaaraha
medium
Duubista Codka
Haa
Afaf
English, Korean, Japanese, Chinese
Noocyada ugu badan
2000

Darwin TTS cod

Default

English
Caadi Neutral

Default (Chinese)

Chinese
Caadi Neutral

Default (Japanese)

Japanese
Caadi Neutral

Default (Korean)

Korean
Caadi Neutral

Darwin TTS Su'aalaha La Weydiiyo

Darwin starts from Qwen3-TTS-1.7B but blends a small fraction of its talker-FFN weights with the matching weights from the Qwen3-1.7B base language model. This training-free blend sharpens cross-lingual voice cloning rather than changing the base voices.

English, Korean, Japanese, and Chinese. The FINAL-Bench release specifically markets its cross-lingual blend for those four, and the deployed model ships voices for them.

About three seconds. It works in zero-shot mode, so no fine-tuning or training is required — you provide a short reference clip and it generates new speech in that voice.
← Codadka oo dhan