GPT-SoVITS TTS
A few-shot voice cloning model that replicates a voice — and can even sing — from as little as five seconds of audio.
Wrap ou tèks nan SSML tags pou presizyon kontwòl:
<speak><prosody rate="slow">Slow speech</prosody></speak>
Tags ke modèl la chwazi konprann — klike pou mete yon nan tèks ou kote li rive:
Modèl sa a li tèks senp, se poutèt sa atik ki nan liy yo pa pran an kont. Pou efè ki baze sou atik, chanje pou yon modèl ekspresyon tankou Orpheus oswa Bark.
Define prononciations Custom (mot = prononciation):
Atik GPT-SoVITS
GPT-SoVITS, created by the developer known as RVC-Boss, combines GPT-style language modeling with SoVITS (Singing Voice Conversion / synthesis) to deliver some of the most accessible voice cloning in open source. With as little as five seconds of reference audio it captures a speaker's timbre and style, and it stands out from most TTS models in handling singing as well as speech. It works across English, Chinese, Japanese, and Korean and supports cross-lingual generation, so a cloned voice can speak a language the reference clip never used. It is widely used by content creators for voice replication, dubbing, and song covers, and reaches high fidelity for a model of its size.
Pi bon pou: Voice cloning, singing synthesis, content creator voice replication
Navigue tout GPT-SoVITS VoyYon ti gade
- Pwogramè
- RVC-Boss
- Lisans
- MIT
- Nivo
- standard
- Vitès
- slow
- Klonaj vwa
- Wi
- Lang
- English, Chinese, Japanese, Korean
- Karakteris maksimòm
- 500