OpenVoice نووسراو بۆ بیستراو
MyShell.ai's instant voice-cloning model with granular control over style, emotion, accent, and rhythm.
نوسراوەکەت بگۆڕە بۆ تگەکانی SSML بۆ کۆنتڕۆڵی ڕاستەقینە:
<speak><prosody rate="slow">Slow speech</prosody></speak>
تگەکان کە مۆدێلی هەڵبژێردراو تێیدەگات - بکەرەوە بۆ ئەوەی یەکێکیان بخەیتە ناو دەقەکەتەوە کە تیایدا ڕوودەدات:
ئەم مۆدێلە نوسراوێکی ئاسایی دەخوێنێتەوە، بۆیە تاگەکانی ناو ڕستە پشتگوێ دەخرێت. بۆ هەستێکی لەسەر بنەمای تاگ، بگەڕێ بۆ مۆدێلێکی دەربڕین وەک ئۆرفیۆس یان بارک.
پێناسەکردنی دەنگی خۆت (وشە = دەنگی):
دەربارەی OpenVoice
OpenVoice from MyShell.ai is built around a ToneColorConverter on top of a MeloTTS backbone, and its defining trait is granular controllability. It clones a voice from a short clip and then lets you independently tune style, emotion, accent, rhythm, pauses, and intonation, generating speech in multiple languages while preserving the speaker's identity. It also doubles as a voice converter for transforming one voice into another. On TTS.ai it covers English, Chinese, Japanese, Korean, French, and Spanish, accepts up to 5,000 characters, and needs roughly 10 seconds of reference audio to clone.
باشترین بۆ: Voice cloning with fine-grained style control, voice conversion
سەردانی هەموویان بکە OpenVoice دەنگیچاوپێکەوتن
- پەرەپێدەر
- MyShell.ai / MIT
- مۆڵەتی بەکارھێنەر
- MIT
- یهمهن
- premium
- خێرایی
- medium
- دووبارە دروستکردنی دەنگی
- بەڵێ
- زمان
- English, Chinese, Japanese, Korean, French, Spanish
- زۆرترین پیت
- 5000