Darwin TTS

Darwin TTS TTS

A Qwen3-TTS variant whose talker FFN weights are blended from the Qwen3 language model for sharper cross-lingual cloning.

Aliĝi for 5, 000 character limit

Envolvu vian tekston en SSML- etikedojn por preciza kontrolo:

<speak><prosody rate="slow">Slow speech</prosody></speak>

Etikedoj kiujn la elektita modelo komprenas - klaku por meti unu en vian tekston kie ĝi okazas:

This model reads plain text, so inline tags are ignored. For tag-based emotion, switch to an expressive model like Orpheus or Bark.

Difini proprajn elparolojn (vorto = elparolo):

-12 +12
0.5x 2.0x
Libera kun Piper, VITS, MeloTTS
Via generita sono aperos tie ĉi. Elektu modelon, entajpu tekston, kaj alklaku Generi.
Sondosiero sukcese generita
0:00
Elŝuti sonon Elŝuti.srt Ligo eksvalidiĝas post 24 horoj
Libera programaro: persona uzo. Komerca licenco ekde $5/mo
Ĉu vi ŝatas TTS.ai? Diru al viaj amikoj!

Pri Darwin TTS

Darwin-TTS-1.7B-Cross by FINAL-Bench is a research variant of Qwen3-TTS-1.7B with an unusual construction: 84 of its talker-FFN tensors (about 8.6% of them) are blended at a 3% ratio with the matching tensors from Qwen3-1.7B-Base, all without any retraining. The result is a model that produces noticeably crisper cross-lingual voice cloning across Korean, English, Japanese, and Chinese — its four core languages. It operates in zero-shot voice-clone mode, needing only about three seconds of reference audio to capture a speaker. Darwin is best suited to transferring a single reference voice across those four languages, for example dubbing or multilingual narration with consistent speaker identity.

Plej bona por: Cross-lingual voice cloning between English / Korean / Japanese / Chinese with a single reference voice

Foliumi ĉiujn Darwin TTS voĉoj

Unu rigardo

Programisto
FINAL-Bench
Licenco
Apache 2.0
Tamuz
standard
Rapideco
medium
Voĉo- klonado
Jes
Lingvoj
English, Korean, Japanese, Chinese
Maksimuma nombro da signoj
2000

Darwin TTS voĉoj

Default

English
Defaŭlta Neutral

Default (Chinese)

Chinese
Defaŭlta Neutral

Default (Japanese)

Japanese
Defaŭlta Neutral

Default (Korean)

Korean
Defaŭlta Neutral

Darwin TTS TTS - FAQ

Darwin starts from Qwen3-TTS-1.7B but blends a small fraction of its talker-FFN weights with the matching weights from the Qwen3-1.7B base language model. This training-free blend sharpens cross-lingual voice cloning rather than changing the base voices.

English, Korean, Japanese, and Chinese. The FINAL-Bench release specifically markets its cross-lingual blend for those four, and the deployed model ships voices for them.

About three seconds. It works in zero-shot mode, so no fine-tuning or training is required — you provide a short reference clip and it generates new speech in that voice.
← Ĉiuj voĉoj