OpenVoice 1 - تكنولوجيا المعلومات والاتصالات
MyShell.ai's instant voice-cloning model with granular control over style, emotion, accent, and rhythm.
لف نصك في علامات SSML للتحكم الدقيق:
<speak><prosody rate="slow">Slow speech</prosody></speak>
العلامات التي يفهمها النموذج المختار — انقر لإسقاط واحدة في نصك حيث تحصل:
هذا النموذج يقرأ النص العادي، لذلك يتم تجاهل العلامات في السطر. للتعبير عن المشاعر القائمة على العلامات، انتقل إلى نموذج تعبيري مثل أورفيوس أو بارك.
تعريف النطق العادي (كلمة = نطق):
حول OpenVoice
OpenVoice from MyShell.ai is built around a ToneColorConverter on top of a MeloTTS backbone, and its defining trait is granular controllability. It clones a voice from a short clip and then lets you independently tune style, emotion, accent, rhythm, pauses, and intonation, generating speech in multiple languages while preserving the speaker's identity. It also doubles as a voice converter for transforming one voice into another. On TTS.ai it covers English, Chinese, Japanese, Korean, French, and Spanish, accepts up to 5,000 characters, and needs roughly 10 seconds of reference audio to clone.
أفضل لل: Voice cloning with fine-grained style control, voice conversion
تصفح جميع OpenVoice الأصواتلمحة عامة
- مطوِّر
- MyShell.ai / MIT
- الترخيص
- MIT
- الرتبة
- premium
- السرعة
- medium
- استنساخ الصوت
- نعم
- اللغات
- English, Chinese, Japanese, Korean, French, Spanish
- الحد الأقصى للحروف
- 5000