OpenVoice ТТС
MyShell.ai's instant voice-cloning model with granular control over style, emotion, accent, and rhythm.
Захоўваць тэкст у тэгах SSML для дакладнага кантролю:
<speak><prosody rate="slow">Slow speech</prosody></speak>
Тэгі, якія разумее выбраная мадэль - націсніце, каб перанесці іх у тэкст:
Гэтая мадэль чытае звычайны тэкст, таму ўбудаваныя тэгі ігнаруюцца. Для эмоцый, заснаваных на тэгах, пераключыцеся на мадэлі выразнасці, такія як Orpheus або Bark.
Вызначыць уласнае вымаўленне (слова = вымаўленне):
Пра OpenVoice
OpenVoice from MyShell.ai is built around a ToneColorConverter on top of a MeloTTS backbone, and its defining trait is granular controllability. It clones a voice from a short clip and then lets you independently tune style, emotion, accent, rhythm, pauses, and intonation, generating speech in multiple languages while preserving the speaker's identity. It also doubles as a voice converter for transforming one voice into another. On TTS.ai it covers English, Chinese, Japanese, Korean, French, and Spanish, accepts up to 5,000 characters, and needs roughly 10 seconds of reference audio to clone.
Лепшы для: Voice cloning with fine-grained style control, voice conversion
Прагляд усіх OpenVoice галасыКароткае апісанне
- Распрацоўшчык
- MyShell.ai / MIT
- Ліцэнзія
- MIT
- Стварыць
- premium
- Хуткасць
- medium
- Клонаванне голасу
- Так
- Мовы
- English, Chinese, Japanese, Korean, French, Spanish
- Найбольшая колькасць знакаў
- 5000