CosyVoice 2

CosyVoice 2 1 - تكنولوجيا المعلومات والاتصالات

Alibaba Tongyi Lab's streaming TTS reaching human-parity naturalness with near-zero latency and zero-shot cloning.

انضم 000 5 كلمة

لف نصك في علامات SSML للتحكم الدقيق:

<speak><prosody rate="slow">Slow speech</prosody></speak>

العلامات التي يفهمها النموذج المختار — انقر لإسقاط واحدة في نصك حيث تحصل:

هذا النموذج يقرأ النص العادي، لذلك يتم تجاهل العلامات في السطر. للتعبير عن المشاعر القائمة على العلامات، انتقل إلى نموذج تعبيري مثل أورفيوس أو بارك.

تعريف النطق العادي (كلمة = نطق):

-12 +12
0.5x 2.0x
مجاني مع Piper, VITS, MeloTTS
سيظهر الصوت الذي أنتجته هنا. اختر نموذجاً، وأدخل نصاً، ثم انقر على توليد.
تم توليد الصوت بنجاح
0:00
تنزيل الصوت تنزيل.srt الرابط ينتهي بعد 24 ساعة
المستوى المجاني: الاستخدام الشخصي. ترخيص تجاري من 5 دولارات شهريا
أحب TTS.ai؟ أخبر أصدقائك!

حول CosyVoice 2

CosyVoice 2, from Alibaba's Tongyi Lab, was designed to make high-quality speech viable in real time. It uses a finite scalar quantization approach combined with flow matching to support streaming synthesis at extremely low latency, while reaching human-comparable naturalness that outperforms many commercial systems in subjective tests. Beyond quality, it offers zero-shot voice cloning from about 3 seconds of audio, cross-lingual synthesis, and fine-grained emotion control. Covering 8 languages with a 1,000-character cap, it's a strong fit for voice assistants, streaming TTS, and other real-time applications.

أفضل لل: Real-time applications, streaming TTS, voice assistants

تصفح جميع CosyVoice 2 الأصوات

لمحة عامة

مطوِّر
Alibaba (Tongyi Lab)
الترخيص
Apache 2.0
الرتبة
standard
السرعة
medium
استنساخ الصوت
نعم
اللغات
English, Chinese, Japanese, Korean, French, German, Italian, Spanish
الحد الأقصى للحروف
1000

CosyVoice 2 الأصوات

Chinese Female

Chinese
المعيار Female

Chinese Male

Chinese
المعيار Male

English Female

English
المعيار Female

English Male

English
المعيار Male

French Female

French
المعيار Female

German Female

German
المعيار Female

Italian Female

Italian
المعيار Female

Japanese Female

Japanese
المعيار Female

Korean Female

Korean
المعيار Female

Spanish Female

Spanish
المعيار Female

CosyVoice 2 الأسئلة المتكررة

Yes. CosyVoice 2 uses finite scalar quantization for streaming synthesis at very low latency, which is what makes it suitable for voice assistants and real-time applications.

Yes. It offers zero-shot voice cloning from roughly 3 seconds of reference audio, plus cross-lingual synthesis and emotion control.

Yes. CosyVoice 2 is Apache 2.0 licensed. It supports 8 languages: English, Chinese, Japanese, Korean, French, German, Italian, and Spanish.
← جميع الأصوات