Kokoro TTS
An 82M-parameter open model from Hexgrad that delivers studio-quality speech at nearly 100x real-time.
Lập vòng văn bản trong thẻ SSML để kiểm soát chính xác:
<speak><prosody rate="slow">Slow speech</prosody></speak>
Thẻ mà mô hình đã chọn hiểu — nhấn để thả một trong văn bản của bạn nơi nó xảy ra:
Mô hình này đọc văn bản đơn giản, vì vậy các thẻ trong dòng sẽ bị bỏ qua. Đối với cảm xúc dựa trên thẻ, hãy chuyển sang mô hình biểu cảm như Orpheus hay Bark.
Định nghĩa cách phát âm tùy chỉnh (từ = phát âm):
Về Kokoro
Kokoro, built by Hexgrad, is a deliberately tiny 82-million-parameter model that punches far above its size class. It uses a StyleTTS + ISTFTNet architecture and was trained on roughly 1,200 hours of speech, yet generates audio close to 100x faster than real-time on a GPU while staying remarkably natural and expressive. It covers English, Japanese, Chinese, French, Italian, Portuguese, Spanish, and Hindi with a varied set of expressive voicepacks, and supports streaming. The combination of small footprint, low latency, and a permissive license has made Kokoro one of the most popular free models for high-volume and streaming use — it carries the largest share of traffic on TTS.ai.
Tốt nhất cho: High-quality TTS with minimal latency, streaming applications
& Xem tất cả Kokoro giọng nóiMột cái nhìn
- Nhà phát triển
- Hexgrad
- Giấy phép
- Apache 2.0
- Thú
- free
- Tốc độ
- fast
- Ký âm
- Không
- Ngôn ngữ
- English, Japanese, Chinese, French, Italian, Portuguese, Spanish, Hindi
- Tối đa các ký tự
- 500