GPT-SoVITS TTS
A few-shot voice cloning model that replicates a voice — and can even sing — from as little as five seconds of audio.
Lập vòng văn bản trong thẻ SSML để kiểm soát chính xác:
<speak><prosody rate="slow">Slow speech</prosody></speak>
Thẻ mà mô hình đã chọn hiểu — nhấn để thả một trong văn bản của bạn nơi nó xảy ra:
Mô hình này đọc văn bản đơn giản, vì vậy các thẻ trong dòng sẽ bị bỏ qua. Đối với cảm xúc dựa trên thẻ, hãy chuyển sang mô hình biểu cảm như Orpheus hay Bark.
Định nghĩa cách phát âm tùy chỉnh (từ = phát âm):
Về GPT-SoVITS
GPT-SoVITS, created by the developer known as RVC-Boss, combines GPT-style language modeling with SoVITS (Singing Voice Conversion / synthesis) to deliver some of the most accessible voice cloning in open source. With as little as five seconds of reference audio it captures a speaker's timbre and style, and it stands out from most TTS models in handling singing as well as speech. It works across English, Chinese, Japanese, and Korean and supports cross-lingual generation, so a cloned voice can speak a language the reference clip never used. It is widely used by content creators for voice replication, dubbing, and song covers, and reaches high fidelity for a model of its size.
Tốt nhất cho: Voice cloning, singing synthesis, content creator voice replication
& Xem tất cả GPT-SoVITS giọng nóiMột cái nhìn
- Nhà phát triển
- RVC-Boss
- Giấy phép
- MIT
- Thú
- standard
- Tốc độ
- slow
- Ký âm
- Có
- Ngôn ngữ
- English, Chinese, Japanese, Korean
- Tối đa các ký tự
- 500