Darwin TTS

Darwin TTS TTS

A Qwen3-TTS variant whose talker FFN weights are blended from the Qwen3 language model for sharper cross-lingual cloning.

Kayıt ol 5000 karakter sınırı için

Kesin kontrol için metninizi SSML etiketleri içinde sarılın:

<speak><prosody rate="slow">Slow speech</prosody></speak>

Seçilmiş modeli anlayan etiketler — bir tanesini metinde olduğu yere bırakmak için tıklayınız:

Bu model sıradan metin okur, bu yüzden satır içindeki etiketler göz ardı edilir. Etiket tabanlı duygular için, Orpheus veya Bark gibi ifadesel bir modele geçin.

Özel telaffuzları tanımla (kelime = telaffuz):

-12 +12
0.5x 2.0x
Piper, VITS, MeloTTS ile ücretsiz
Oluşturduğunuz ses burada görünecek. Bir model seçin, metni girin ve Oluştur' a basın.
Ses Başarıyla Oluşturuldu
0:00
Ses İndir İndir Bağlantı 24 saat içinde sona erer
Özgür katman: kişisel kullanım. Ticari lisans ayda 5$'dan
TTS.ai'yi seviyor musunuz?

Hakkında Darwin TTS

Darwin-TTS-1.7B-Cross by FINAL-Bench is a research variant of Qwen3-TTS-1.7B with an unusual construction: 84 of its talker-FFN tensors (about 8.6% of them) are blended at a 3% ratio with the matching tensors from Qwen3-1.7B-Base, all without any retraining. The result is a model that produces noticeably crisper cross-lingual voice cloning across Korean, English, Japanese, and Chinese — its four core languages. It operates in zero-shot voice-clone mode, needing only about three seconds of reference audio to capture a speaker. Darwin is best suited to transferring a single reference voice across those four languages, for example dubbing or multilingual narration with consistent speaker identity.

En iyi: Cross-lingual voice cloning between English / Korean / Japanese / Chinese with a single reference voice

Hepsini Ara Darwin TTS sesleri

Bir bakışta

Geliştirici
FINAL-Bench
Lisans
Apache 2.0
Hayvan
standard
Hız
medium
Ses klonlama
Evet
Dilleri
English, Korean, Japanese, Chinese
Maksimum karakterler
2000

Darwin TTS sesleri

Default

English
Standart Neutral

Default (Chinese)

Chinese
Standart Neutral

Default (Japanese)

Japanese
Standart Neutral

Default (Korean)

Korean
Standart Neutral

Darwin TTS TTS - FAQ

Darwin starts from Qwen3-TTS-1.7B but blends a small fraction of its talker-FFN weights with the matching weights from the Qwen3-1.7B base language model. This training-free blend sharpens cross-lingual voice cloning rather than changing the base voices.

English, Korean, Japanese, and Chinese. The FINAL-Bench release specifically markets its cross-lingual blend for those four, and the deployed model ships voices for them.

About three seconds. It works in zero-shot mode, so no fine-tuning or training is required — you provide a short reference clip and it generates new speech in that voice.
← Tüm sesleri