OpenVoice 音声翻訳
MyShell.ai's instant voice-cloning model with granular control over style, emotion, accent, and rhythm.
SSML タグでテキストを囲み、正確な制御を行う:
<speak><prosody rate="slow">Slow speech</prosody></speak>
選択したモデルが理解するタグ - クリックしてテキストにドラッグします:
このモデルは単純テキストを読み込み、インラインタグは無視されます。タグベースの感情を表現するには、Orpheus や Bark のような表現モデルに切り替えてください。
カスタム発音を定義 (単語=発音):
情報 OpenVoice
OpenVoice from MyShell.ai is built around a ToneColorConverter on top of a MeloTTS backbone, and its defining trait is granular controllability. It clones a voice from a short clip and then lets you independently tune style, emotion, accent, rhythm, pauses, and intonation, generating speech in multiple languages while preserving the speaker's identity. It also doubles as a voice converter for transforming one voice into another. On TTS.ai it covers English, Chinese, Japanese, Korean, French, and Spanish, accepts up to 5,000 characters, and needs roughly 10 seconds of reference audio to clone.
適合する: Voice cloning with fine-grained style control, voice conversion
すべてブラウズ OpenVoice 声概要
- 開発者
- MyShell.ai / MIT
- ライセンス
- MIT
- 動物
- premium
- スピード
- medium
- 声のクローン
- はい
- 言語
- English, Chinese, Japanese, Korean, French, Spanish
- 最大文字数
- 5000