GPT-SoVITS เสียง
A few-shot voice cloning model that replicates a voice — and can even sing — from as little as five seconds of audio.
หมุนข้อความของคุณในแท็ก SSML เพื่อควบคุมอย่างแม่นยำ:
<speak><prosody rate="slow">Slow speech</prosody></speak>
แท็กที่โมเดลที่เลือกไว้เข้าใจ - คลิกเพื่อวางแท็กลงในข้อความของคุณที่มันเกิดขึ้น:
โมเดลนี้อ่านข้อความธรรมดา ดังนั้น แท็กในบรรทัดจะถูกละเลย สำหรับอารมณ์ที่ใช้แท็ก เปลี่ยนไปใช้โมเดลแสดงออก เช่น Orpheus หรือ Bark
ตั้งค่าการออกเสียงที่กำหนดเอง (คำ = การออกเสียง):
เกี่ยวกับ GPT-SoVITS
GPT-SoVITS, created by the developer known as RVC-Boss, combines GPT-style language modeling with SoVITS (Singing Voice Conversion / synthesis) to deliver some of the most accessible voice cloning in open source. With as little as five seconds of reference audio it captures a speaker's timbre and style, and it stands out from most TTS models in handling singing as well as speech. It works across English, Chinese, Japanese, and Korean and supports cross-lingual generation, so a cloned voice can speak a language the reference clip never used. It is widely used by content creators for voice replication, dubbing, and song covers, and reaches high fidelity for a model of its size.
เหมาะสำหรับ: Voice cloning, singing synthesis, content creator voice replication
แสดงทั้งหมด GPT-SoVITS เสียงเพียงแค่มองดู
- ผู้พัฒนา
- RVC-Boss
- ใบอนุญาต
- MIT
- สัตว์
- standard
- ความเร็ว
- slow
- เสียง
- ใช่
- ภาษา
- English, Chinese, Japanese, Korean
- จำนวนตัวอักษรสูงสุด
- 500