Kokoro TTS
An 82M-parameter open model from Hexgrad that delivers studio-quality speech at nearly 100x real-time.
Întoarceți textul în etichetele SSML pentru un control precis:
<speak><prosody rate="slow">Slow speech</prosody></speak>
Etichetele modelului selectat înțeleg — click pentru a lăsa unul în textul tău unde se întâmplă:
Acest model citește textul simplu, astfel încât etichetele inline sunt ignorate. Pentru emoții bazate pe tag, schimbați la un model expresiv cum ar fi Orpheus sau Bark.
Definiți pronunțiare personalizată (cuvânt = pronunție):
Despre Kokoro
Kokoro, built by Hexgrad, is a deliberately tiny 82-million-parameter model that punches far above its size class. It uses a StyleTTS + ISTFTNet architecture and was trained on roughly 1,200 hours of speech, yet generates audio close to 100x faster than real-time on a GPU while staying remarkably natural and expressive. It covers English, Japanese, Chinese, French, Italian, Portuguese, Spanish, and Hindi with a varied set of expressive voicepacks, and supports streaming. The combination of small footprint, low latency, and a permissive license has made Kokoro one of the most popular free models for high-volume and streaming use — it carries the largest share of traffic on TTS.ai.
Cel mai bun pentru: High-quality TTS with minimal latency, streaming applications
Navigați toate Kokoro vociLa o privire
- Dezvoltator
- Hexgrad
- Licență
- Apache 2.0
- Nivel
- free
- Viteză
- fast
- Clonarea vocală
- Nu.
- Limbi
- English, Japanese, Chinese, French, Italian, Portuguese, Spanish, Hindi
- Caractere maxime
- 500