OpenVoice ТТС
MyShell.ai's instant voice-cloning model with granular control over style, emotion, accent, and rhythm.
Заверните текст в SSML для точного контроля:
<speak><prosody rate="slow">Slow speech</prosody></speak>
Помечает выбранную модель, которая понимает — нажмите, чтобы выкинуть одну в ваш текст, где она случается:
Эта модель читает простой текст, поэтому встраиваемые метки игнорируются. Для эмоций на основе метки переключайтесь на экспрессивную модель, как Орфей или Барк.
Определить традиционные произношения (слово = произношение):
О том, что OpenVoice
OpenVoice from MyShell.ai is built around a ToneColorConverter on top of a MeloTTS backbone, and its defining trait is granular controllability. It clones a voice from a short clip and then lets you independently tune style, emotion, accent, rhythm, pauses, and intonation, generating speech in multiple languages while preserving the speaker's identity. It also doubles as a voice converter for transforming one voice into another. On TTS.ai it covers English, Chinese, Japanese, Korean, French, and Spanish, accepts up to 5,000 characters, and needs roughly 10 seconds of reference audio to clone.
Лучший для: Voice cloning with fine-grained style control, voice conversion
Просмотр OpenVoice голосаВзгляните.
- Разработчик
- MyShell.ai / MIT
- Лицензия
- MIT
- Тяжелый
- premium
- Скорость
- medium
- Клонирование голоса
- Выполнено
- Знание языков
- English, Chinese, Japanese, Korean, French, Spanish
- Максимум символов
- 5000