Darwin TTS

Darwin TTS TTS

A Qwen3-TTS variant whose talker FFN weights are blended from the Qwen3 language model for sharper cross-lingual cloning.

Iscriviti per un limite di 5.000 caratteri

Avvolgi il tuo testo nei tag SSML per un controllo preciso:

<speak><prosody rate="slow">Slow speech</prosody></speak>

Tags il modello selezionato comprende clic su

Questo modello legge testo semplice, quindi i tag inline vengono ignorati. Per le emozioni basate sui tag, passare a un modello espressivo come Orpheus o Bark.

Definire le pronunciazioni personalizzate (parola = pronuncia):

-12 +12
0.5x 2.0x
Gratis con Piper, VITS, MeloTTS
L'audio generato apparirà qui. Scegli un modello, inserisci testo e fai clic su Genera.
Audio generato con successo
0:00
Scarica audio Scarica.srt Link scade in 24 ore
Livello libero: uso personale. Licenza commerciale da $5/mo
Fai di questo la tua voce Clona una voce in 30 secondi
Ti piace TTS.ai? Dillo ai tuoi amici!

Informazioni Darwin TTS

Darwin-TTS-1.7B-Cross by FINAL-Bench is a research variant of Qwen3-TTS-1.7B with an unusual construction: 84 of its talker-FFN tensors (about 8.6% of them) are blended at a 3% ratio with the matching tensors from Qwen3-1.7B-Base, all without any retraining. The result is a model that produces noticeably crisper cross-lingual voice cloning across Korean, English, Japanese, and Chinese — its four core languages. It operates in zero-shot voice-clone mode, needing only about three seconds of reference audio to capture a speaker. Darwin is best suited to transferring a single reference voice across those four languages, for example dubbing or multilingual narration with consistent speaker identity.

Meglio per: Cross-lingual voice cloning between English / Korean / Japanese / Chinese with a single reference voice

Sfoglia tutti Darwin TTS voci

A colpo d'occhio

Sviluppatore
FINAL-Bench
Licenza
Apache 2.0
Livello
standard
Velocità
medium
Clonazione vocale
Lingue
English, Korean, Japanese, Chinese
Caratteri massimi
2000

Darwin TTS voci

Default

English
Standard Neutral

Default (Chinese)

Chinese
Standard Neutral

Default (Japanese)

Japanese
Standard Neutral

Default (Korean)

Korean
Standard Neutral

Darwin TTS FAQ del TTS

Darwin starts from Qwen3-TTS-1.7B but blends a small fraction of its talker-FFN weights with the matching weights from the Qwen3-1.7B base language model. This training-free blend sharpens cross-lingual voice cloning rather than changing the base voices.

English, Korean, Japanese, and Chinese. The FINAL-Bench release specifically markets its cross-lingual blend for those four, and the deployed model ships voices for them.

About three seconds. It works in zero-shot mode, so no fine-tuning or training is required — you provide a short reference clip and it generates new speech in that voice.
← Tutte le voci