OpenVoice TTS
MyShell.ai's instant voice-cloning model with granular control over style, emotion, accent, and rhythm.
Kpọchie ngwe gị n'ime SSML táàbụ̀ maka nlekọta ziri ezi:
<speak><prosody rate="slow">Slow speech</prosody></speak>
Táàbụ̀ nke móòdù ahụ a họọrọ na-aghọta - pịa ka ịkpụga otu n'ime ngwe gị ebe ọ na-eme:
Móòdù a na-agụ ngwe nkịtị, yabụ na a na-ewepụta inline táàbụ̀. Maka táàbụ̀-n'okpuru n'émóòdù, gbanwee ka móòdù na-egosi ihe dịka Orpheus mọọbụ Bark.
Ndesịta okwu emeredịkachọrọ:
_N'ihe banyere OpenVoice
OpenVoice from MyShell.ai is built around a ToneColorConverter on top of a MeloTTS backbone, and its defining trait is granular controllability. It clones a voice from a short clip and then lets you independently tune style, emotion, accent, rhythm, pauses, and intonation, generating speech in multiple languages while preserving the speaker's identity. It also doubles as a voice converter for transforming one voice into another. On TTS.ai it covers English, Chinese, Japanese, Korean, French, and Spanish, accepts up to 5,000 characters, and needs roughly 10 seconds of reference audio to clone.
Ọkachasị maka: Voice cloning with fine-grained style control, voice conversion
Nlegharịa niile OpenVoice ụdaN'ime nlele
- Ńkwádò
- MyShell.ai / MIT
- Ikikere
- MIT
- Tier
- premium
- Nhazi
- medium
- Nhazi ụda
- Ee
- Asụsụ ndị ahụ
- English, Chinese, Japanese, Korean, French, Spanish
- Ụhara Max
- 5000