OpenVoice TTS
MyShell.ai's instant voice-cloning model with granular control over style, emotion, accent, and rhythm.
Envolvu vian tekston en SSML- etikedojn por preciza kontrolo:
<speak><prosody rate="slow">Slow speech</prosody></speak>
Etikedoj kiujn la elektita modelo komprenas - klaku por meti unu en vian tekston kie ĝi okazas:
This model reads plain text, so inline tags are ignored. For tag-based emotion, switch to an expressive model like Orpheus or Bark.
Difini proprajn elparolojn (vorto = elparolo):
Pri OpenVoice
OpenVoice from MyShell.ai is built around a ToneColorConverter on top of a MeloTTS backbone, and its defining trait is granular controllability. It clones a voice from a short clip and then lets you independently tune style, emotion, accent, rhythm, pauses, and intonation, generating speech in multiple languages while preserving the speaker's identity. It also doubles as a voice converter for transforming one voice into another. On TTS.ai it covers English, Chinese, Japanese, Korean, French, and Spanish, accepts up to 5,000 characters, and needs roughly 10 seconds of reference audio to clone.
Plej bona por: Voice cloning with fine-grained style control, voice conversion
Foliumi ĉiujn OpenVoice voĉojUnu rigardo
- Programisto
- MyShell.ai / MIT
- Licenco
- MIT
- Tamuz
- premium
- Rapideco
- medium
- Voĉo- klonado
- Jes
- Lingvoj
- English, Chinese, Japanese, Korean, French, Spanish
- Maksimuma nombro da signoj
- 5000