Бусад
Darwin TTS

Darwin TTS ТТС

A Qwen3-TTS variant whose talker FFN weights are blended from the Qwen3 language model for sharper cross-lingual cloning.

Бүртгүүлэх 5000 тэмдэгтээс хэтрэхгүй

Тодорхой хяналтын тулд SSML тэмдгээр текстээ буулгах:

<speak><prosody rate="slow">Slow speech</prosody></speak>

Бүх

Энэ загвар нь энгийн утга уншдаг, тиймээс доторх тэмдгийг үл тоомсорлодог. Тэг- суурилсан сэтгэл хөдлөл, Orpheus эсвэл Bark- шиг илэрхийлэх загвар руу шилжинэ.

Өөрийн дуудлагыг тодорхойлох (үг = дуудлага):

-12 +12
0.5x 2.0x
Piper, VITS, MeloTTS-тэй чөлөөт
Таны үүсгэсэн дууны файл энд гарч ирнэ. Модель сонгож, текстийг оруулж, Бүтээгдэх товчийг дарна уу.
Аудио амжилттай бүтээгдсэн
0:00
Дуу татаж авах .srt татаж авах Холбоо 24 цагийн дараа дуусна
Хязгааргүй: хувийн хэрэглээ. Бизнесийн лиценз $5/сараас
TTS.ai-г хайрладаг уу? Найзуудаа хэлж өгөөрэй!

Тодорхойлолт Darwin TTS

Darwin-TTS-1.7B-Cross by FINAL-Bench is a research variant of Qwen3-TTS-1.7B with an unusual construction: 84 of its talker-FFN tensors (about 8.6% of them) are blended at a 3% ratio with the matching tensors from Qwen3-1.7B-Base, all without any retraining. The result is a model that produces noticeably crisper cross-lingual voice cloning across Korean, English, Japanese, and Chinese — its four core languages. It operates in zero-shot voice-clone mode, needing only about three seconds of reference audio to capture a speaker. Darwin is best suited to transferring a single reference voice across those four languages, for example dubbing or multilingual narration with consistent speaker identity.

Хамгийн сайн: Cross-lingual voice cloning between English / Korean / Japanese / Chinese with a single reference voice

Бүхнийг харах Darwin TTS дуунууд

Нүүр хуудас

Хөгжүүлэгч
FINAL-Bench
Лиценз
Apache 2.0
Үхрийн
standard
Хурд
medium
Дууны дугуй
Тийм
Хэл
English, Korean, Japanese, Chinese
Хамгийн их үсэг
2000

Darwin TTS дуунууд

Default

English
Стандарт Neutral

Default (Chinese)

Chinese
Стандарт Neutral

Default (Japanese)

Japanese
Стандарт Neutral

Default (Korean)

Korean
Стандарт Neutral

Darwin TTS ТТС - Тодорхойгүй асуултууд

Darwin starts from Qwen3-TTS-1.7B but blends a small fraction of its talker-FFN weights with the matching weights from the Qwen3-1.7B base language model. This training-free blend sharpens cross-lingual voice cloning rather than changing the base voices.

English, Korean, Japanese, and Chinese. The FINAL-Bench release specifically markets its cross-lingual blend for those four, and the deployed model ships voices for them.

About three seconds. It works in zero-shot mode, so no fine-tuning or training is required — you provide a short reference clip and it generates new speech in that voice.
← Бүх дуунууд