GPT-SoVITS TTS
A few-shot voice cloning model that replicates a voice — and can even sing — from as little as five seconds of audio.
Wrap uw tekst in SSML-tags voor nauwkeurige controle:
<speak><prosody rate="slow">Slow speech</prosody></speak>
Tags het geselecteerde model begrijpt
Dit model leest platte tekst, dus inline tags worden genegeerd. Voor emotie op basis van tags, schakel naar een expressief model zoals Orpheus of Bark.
Definieer aangepaste uitspraaken (woord = uitspraak):
Info GPT-SoVITS
GPT-SoVITS, created by the developer known as RVC-Boss, combines GPT-style language modeling with SoVITS (Singing Voice Conversion / synthesis) to deliver some of the most accessible voice cloning in open source. With as little as five seconds of reference audio it captures a speaker's timbre and style, and it stands out from most TTS models in handling singing as well as speech. It works across English, Chinese, Japanese, and Korean and supports cross-lingual generation, so a cloned voice can speak a language the reference clip never used. It is widely used by content creators for voice replication, dubbing, and song covers, and reaches high fidelity for a model of its size.
Beste voor: Voice cloning, singing synthesis, content creator voice replication
Alles doorbladeren GPT-SoVITS stemmenIn een oogopslag
- Ontwikkelaar
- RVC-Boss
- Licentie
- MIT
- Niveau
- standard
- Snelheid
- slow
- Klonen van stemmen
- Ja.
- Talen
- English, Chinese, Japanese, Korean
- Max. tekens
- 500