OpenVoice TTS
MyShell.ai's instant voice-cloning model with granular control over style, emotion, accent, and rhythm.
Ajusta el text a les etiquetes SSML per al control precís:
<speak><prosody rate="slow">Slow speech</prosody></speak>
Etiquetes del model seleccionat entenen el clic show clic per a deixar- ne un al text a on succeeix:
Aquest model llegeix text pla, així que les etiquetes inserides s' ignoren. Per a emocions basades en etiquetes, canvieu a un model expressiu com Orfeus o Bark.
Defineix pronúncies personalitzades (word = pronunciació):
Quant a OpenVoice
OpenVoice from MyShell.ai is built around a ToneColorConverter on top of a MeloTTS backbone, and its defining trait is granular controllability. It clones a voice from a short clip and then lets you independently tune style, emotion, accent, rhythm, pauses, and intonation, generating speech in multiple languages while preserving the speaker's identity. It also doubles as a voice converter for transforming one voice into another. On TTS.ai it covers English, Chinese, Japanese, Korean, French, and Spanish, accepts up to 5,000 characters, and needs roughly 10 seconds of reference audio to clone.
Millor per: Voice cloning with fine-grained style control, voice conversion
Navega- ho tot OpenVoice veusEn una mirada
- Desenvolupador
- MyShell.ai / MIT
- Llicència
- MIT
- TierCity name (optional, probably does not need a translation)
- premium
- Velocitat
- medium
- clonació de veu
- Sí
- Idiomes
- English, Chinese, Japanese, Korean, French, Spanish
- Nombre màxim de caràcters
- 5000