Kokoro เสียง
An 82M-parameter open model from Hexgrad that delivers studio-quality speech at nearly 100x real-time.
หมุนข้อความของคุณในแท็ก SSML เพื่อควบคุมอย่างแม่นยำ:
<speak><prosody rate="slow">Slow speech</prosody></speak>
แท็กที่โมเดลที่เลือกไว้เข้าใจ - คลิกเพื่อวางแท็กลงในข้อความของคุณที่มันเกิดขึ้น:
โมเดลนี้อ่านข้อความธรรมดา ดังนั้น แท็กในบรรทัดจะถูกละเลย สำหรับอารมณ์ที่ใช้แท็ก เปลี่ยนไปใช้โมเดลแสดงออก เช่น Orpheus หรือ Bark
ตั้งค่าการออกเสียงที่กำหนดเอง (คำ = การออกเสียง):
เกี่ยวกับ Kokoro
Kokoro, built by Hexgrad, is a deliberately tiny 82-million-parameter model that punches far above its size class. It uses a StyleTTS + ISTFTNet architecture and was trained on roughly 1,200 hours of speech, yet generates audio close to 100x faster than real-time on a GPU while staying remarkably natural and expressive. It covers English, Japanese, Chinese, French, Italian, Portuguese, Spanish, and Hindi with a varied set of expressive voicepacks, and supports streaming. The combination of small footprint, low latency, and a permissive license has made Kokoro one of the most popular free models for high-volume and streaming use — it carries the largest share of traffic on TTS.ai.
เหมาะสำหรับ: High-quality TTS with minimal latency, streaming applications
แสดงทั้งหมด Kokoro เสียงเพียงแค่มองดู
- ผู้พัฒนา
- Hexgrad
- ใบอนุญาต
- Apache 2.0
- สัตว์
- free
- ความเร็ว
- fast
- เสียง
- ไม่มี
- ภาษา
- English, Japanese, Chinese, French, Italian, Portuguese, Spanish, Hindi
- จำนวนตัวอักษรสูงสุด
- 500