OpenVoice TTS
MyShell.ai's instant voice-cloning model with granular control over style, emotion, accent, and rhythm.
Wrap your text in SSML tags for precise control:
<speak><prosody rate="slow">Slow speech</prosody></speak>
Tags déi d'gewielt Modell verstinn - klickt fir eng an Ärem Text ze setzen wou se geschitt:
Dëse Modell liest einfache Text, sou datt Inline-Tags ignoréiert ginn. Fir Tag-baséiert Emotiounen, wielt e expressiven Modell wéi Orpheus oder Bark.
Eegen Aussproochen definéieren (Wuert = Aussprooch):
Iwwer OpenVoice
OpenVoice from MyShell.ai is built around a ToneColorConverter on top of a MeloTTS backbone, and its defining trait is granular controllability. It clones a voice from a short clip and then lets you independently tune style, emotion, accent, rhythm, pauses, and intonation, generating speech in multiple languages while preserving the speaker's identity. It also doubles as a voice converter for transforming one voice into another. On TTS.ai it covers English, Chinese, Japanese, Korean, French, and Spanish, accepts up to 5,000 characters, and needs roughly 10 seconds of reference audio to clone.
Bescht fir: Voice cloning with fine-grained style control, voice conversion
All sichen OpenVoice StimmenOp ee Bléck
- Entwéckler
- MyShell.ai / MIT
- Lizenz
- MIT
- Tier
- premium
- Geschwindegkeet
- medium
- Sprooche-Klonen
- Ja
- Sproochen
- English, Chinese, Japanese, Korean, French, Spanish
- Maximal Zeichen
- 5000