OpenVoice TTS
MyShell.ai's instant voice-cloning model with granular control over style, emotion, accent, and rhythm.
Wrap test tiegħek fil-tags SSML għall-kontroll preċiż:
<speak><prosody rate="slow">Slow speech</prosody></speak>
Tags li l-mudell magħżul jifhem — ikklikkja biex tqiegħed waħda fit-test tiegħek fejn jiġri:
Dan il-mudell jaqra test sempliċi, għalhekk it-tags inline huma injorati.Għal emozzjoni bbażata fuq it-tag, aqleb għal mudell espressiv bħal Orpheus jew Bark.
Iddefinixxi pronunzji tad-dwana (kelma = pronunzja):
Dwar OpenVoice
OpenVoice from MyShell.ai is built around a ToneColorConverter on top of a MeloTTS backbone, and its defining trait is granular controllability. It clones a voice from a short clip and then lets you independently tune style, emotion, accent, rhythm, pauses, and intonation, generating speech in multiple languages while preserving the speaker's identity. It also doubles as a voice converter for transforming one voice into another. On TTS.ai it covers English, Chinese, Japanese, Korean, French, and Spanish, accepts up to 5,000 characters, and needs roughly 10 seconds of reference audio to clone.
L-aħjar għal: Voice cloning with fine-grained style control, voice conversion
Ibbrawżja kollox OpenVoice vuċijietDaqqa t'għajn
- Żviluppatur
- MyShell.ai / MIT
- Liċenzja
- MIT
- Annimali
- premium
- Veloċità
- medium
- Klonazzjoni tal-vuċi
- Iva
- Lingwi
- English, Chinese, Japanese, Korean, French, Spanish
- Karattri massimi
- 5000