Kokoro TTS
An 82M-parameter open model from Hexgrad that delivers studio-quality speech at nearly 100x real-time.
Verpacken Sie Ihren Text in SSML-Tags für eine präzise Kontrolle:
<speak><prosody rate="slow">Slow speech</prosody></speak>
Tags, die das ausgewählte Modell versteht — klicken Sie, um einen in Ihren Text zu legen, wo es passiert:
Dieses Modell liest Text, so dass Inline-Tags ignoriert werden. Für tag-basierte Emotion, wechseln Sie zu einem ausdrucksstarken Modell wie Orpheus oder Bark.
Benutzerdefinierte Aussprachen definieren (Wort = Aussprache):
Über Kokoro
Kokoro, built by Hexgrad, is a deliberately tiny 82-million-parameter model that punches far above its size class. It uses a StyleTTS + ISTFTNet architecture and was trained on roughly 1,200 hours of speech, yet generates audio close to 100x faster than real-time on a GPU while staying remarkably natural and expressive. It covers English, Japanese, Chinese, French, Italian, Portuguese, Spanish, and Hindi with a varied set of expressive voicepacks, and supports streaming. The combination of small footprint, low latency, and a permissive license has made Kokoro one of the most popular free models for high-volume and streaming use — it carries the largest share of traffic on TTS.ai.
Das Beste für: High-quality TTS with minimal latency, streaming applications
Alle durchsuchen Kokoro StimmenAuf einen Blick
- Entwickler
- Hexgrad
- Lizenz
- Apache 2.0
- Tierart
- free
- Geschwindigkeit
- fast
- Klonen der Stimme
- Nein
- Sprachen
- English, Japanese, Chinese, French, Italian, Portuguese, Spanish, Hindi
- Maximale Zeichen
- 500