OpenVoice TTS
MyShell.ai's instant voice-cloning model with granular control over style, emotion, accent, and rhythm.
برای کنترل دقیق ، متن خود را در برچسبهای SSML بپیچید:
<speak><prosody rate="slow">Slow speech</prosody></speak>
برچسبهایی که مدل برگزیده میفهمد — برای انداختن یکی در متن خود ، جایی که اتفاق میافتد ، کلیک کنید:
این مدل متن ساده را میخواند ، بنابراین برچسبهای خطی نادیده گرفته میشوند. برای احساسات مبتنی بر برچسب ، به یک مدل بیانی مانند Orpheus یا Bark تغییر دهید.
تعریف تلفظ سفارشی) کلمه = تلفظ (:
در مورد OpenVoice
OpenVoice from MyShell.ai is built around a ToneColorConverter on top of a MeloTTS backbone, and its defining trait is granular controllability. It clones a voice from a short clip and then lets you independently tune style, emotion, accent, rhythm, pauses, and intonation, generating speech in multiple languages while preserving the speaker's identity. It also doubles as a voice converter for transforming one voice into another. On TTS.ai it covers English, Chinese, Japanese, Korean, French, and Spanish, accepts up to 5,000 characters, and needs roughly 10 seconds of reference audio to clone.
بهترین برای: Voice cloning with fine-grained style control, voice conversion
مرور همۀ OpenVoice صداهايه نگاهي بنداز
- توسعهدهنده
- MyShell.ai / MIT
- مجوز
- MIT
- حیوان
- premium
- سرعت
- medium
- شبیهسازی صدا
- آره
- زبانها
- English, Chinese, Japanese, Korean, French, Spanish
- بیشینه نویسهها
- 5000