Darwin TTS

Darwin TTS TTS

A Qwen3-TTS variant whose talker FFN weights are blended from the Qwen3 language model for sharper cross-lingual cloning.

Registreeru 5000 tähemärgi piir

SSML-i siltidesse teksti segamine täpseks kontrollimiseks:

<speak><prosody rate="slow">Slow speech</prosody></speak>

Sildid valitud mudelil mõistavad ~ klõpsa ühe kukutamiseks teksti, kus see juhtub:

See mudel loeb lihtsat teksti, nii et sisemisi silte ignoreeritakse. Sildil põhinevate emotsioonide puhul lülituge ekspressiivsele mudelile nagu Orpheus või Bark.

Kohandatud häälduste määramine (sõna = hääldus):

-12 +12
0.5x 2.0x
Tasuta Piper, VITS, MeloTTS
Siin ilmub sinu loodud heli. Vali mudel, sisesta tekst ja klõpsa Genereeri.
Audio genereeritud edukalt
0:00
Audio allalaadimine Lae alla.srt Link aegub 24 tunni pärast.
Tasuta tase: isiklik kasutamine. Äriline litsents alates $5/mo
Armastus TTS.ai?

Info Darwin TTS

Darwin-TTS-1.7B-Cross by FINAL-Bench is a research variant of Qwen3-TTS-1.7B with an unusual construction: 84 of its talker-FFN tensors (about 8.6% of them) are blended at a 3% ratio with the matching tensors from Qwen3-1.7B-Base, all without any retraining. The result is a model that produces noticeably crisper cross-lingual voice cloning across Korean, English, Japanese, and Chinese — its four core languages. It operates in zero-shot voice-clone mode, needing only about three seconds of reference audio to capture a speaker. Darwin is best suited to transferring a single reference voice across those four languages, for example dubbing or multilingual narration with consistent speaker identity.

Parim: Cross-lingual voice cloning between English / Korean / Japanese / Chinese with a single reference voice

Kõigi sirvimine Darwin TTS hääled

Põgusalt

Arendaja
FINAL-Bench
Litsents
Apache 2.0
Määramistasand
standard
Kiirus
medium
Hääle kloonimine
Jah
Keeled
English, Korean, Japanese, Chinese
Maks. märgid
2000

Darwin TTS hääled

Default

English
Standardne Neutral

Default (Chinese)

Chinese
Standardne Neutral

Default (Japanese)

Japanese
Standardne Neutral

Default (Korean)

Korean
Standardne Neutral

Darwin TTS TTS (KKK)

Darwin starts from Qwen3-TTS-1.7B but blends a small fraction of its talker-FFN weights with the matching weights from the Qwen3-1.7B base language model. This training-free blend sharpens cross-lingual voice cloning rather than changing the base voices.

English, Korean, Japanese, and Chinese. The FINAL-Bench release specifically markets its cross-lingual blend for those four, and the deployed model ships voices for them.

About three seconds. It works in zero-shot mode, so no fine-tuning or training is required — you provide a short reference clip and it generates new speech in that voice.
← Kõik hääled