OpenVoice TTS
MyShell.ai's instant voice-cloning model with granular control over style, emotion, accent, and rhythm.
Ukufaka umbhalo wakho kumathegi we-SSML ukulawula okucacile:
<speak><prosody rate="slow">Slow speech</prosody></speak>
Amathegi amamodeli akhethiwe aqonda - chofoza ukuwasusa kusihloko sakho lapho kwenzeka khona:
Le modeli ifunda umbhalo ojwayelekile, ngakho amathegi e-inline akhohlwa. Ukwenza umbono osekelwe kumathegi, shintsha kwimodeli ebonisa umbono njenge-Orpheus noma i-Bark.
Chaza ukuchaza okujwayelekile (igama = ukuchaza):
Ngo OpenVoice
OpenVoice from MyShell.ai is built around a ToneColorConverter on top of a MeloTTS backbone, and its defining trait is granular controllability. It clones a voice from a short clip and then lets you independently tune style, emotion, accent, rhythm, pauses, and intonation, generating speech in multiple languages while preserving the speaker's identity. It also doubles as a voice converter for transforming one voice into another. On TTS.ai it covers English, Chinese, Japanese, Korean, French, and Spanish, accepts up to 5,000 characters, and needs roughly 10 seconds of reference audio to clone.
Okungcono kakhulu: Voice cloning with fine-grained style control, voice conversion
Khangela konke OpenVoice izizwiNgombono ocacile
- Umthuthukisi
- MyShell.ai / MIT
- Ilayisense
- MIT
- I-Tiger
- premium
- Isivinini
- medium
- Ukuklona umsindo
- Yebo
- Izilimi
- English, Chinese, Japanese, Korean, French, Spanish
- Amaphawu aphezulu
- 5000