Kokoro

Kokoro TTS

An 82M-parameter open model from Hexgrad that delivers studio-quality speech at nearly 100x real-time.

Registrer deg for 5000 tegn- grense

Bryt teksten i SSML- tagger for nøyaktig kontroll:

<speak><prosody rate="slow">Slow speech</prosody></speak>

Merker den valgte modellen forstår – klikk for å slippe en i teksten der den skjer:

Denne modellen leser ren tekst, så merker i teksten blir ignorert. Bytt til en ekspressiv modell som Orfeus eller Bark for å bruke tagger.

Definer selvvalgte uttaler (ord = uttale):

-12 +12
0.5x 2.0x
Fri for piper, VITS, MeloTTS
Her vises din genererte lyd. Velg en modell, skriv inn tekst og trykk Generer.
Lydgenerert vellykket
0:00
Last ned lyd Last ned.srt Lenke utløper om 24 timer
Fritt nivå: personlig bruk. Handelslisens fra $5/mo
Elsker TTS.ai? Fortell vennene dine!

Om Kokoro

Kokoro, built by Hexgrad, is a deliberately tiny 82-million-parameter model that punches far above its size class. It uses a StyleTTS + ISTFTNet architecture and was trained on roughly 1,200 hours of speech, yet generates audio close to 100x faster than real-time on a GPU while staying remarkably natural and expressive. It covers English, Japanese, Chinese, French, Italian, Portuguese, Spanish, and Hindi with a varied set of expressive voicepacks, and supports streaming. The combination of small footprint, low latency, and a permissive license has made Kokoro one of the most popular free models for high-volume and streaming use — it carries the largest share of traffic on TTS.ai.

Best for: High-quality TTS with minimal latency, streaming applications

Bla gjennom alle Kokoro stemmer

Med et blikk

Utvikler
Hexgrad
Lisens
Apache 2.0
Nivå
free
Hastighet
fast
Stemmekloning
Nei
Språk
English, Japanese, Chinese, French, Italian, Portuguese, Spanish, Hindi
Største antall tegn
500

Kokoro stemmer

Adam

English
Ledig Male

Alex

Spanish
Ledig Male

Alex

Portuguese
Ledig Male

Alpha

Japanese
Ledig Female

Alpha

Hindi
Ledig Female

Bella

English
Ledig Female

Dora

Portuguese
Ledig Female

Dora

Spanish
Ledig Female

Emma (British)

English
Ledig Female

George (British)

English
Ledig Male

Gongitsune

Japanese
Ledig Female

Heart

English
Ledig Female

Isabella (British)

English
Ledig Female

Lewis (British)

English
Ledig Male

Michael

English
Ledig Male

Nicola

Italian
Ledig Male

Nicole

English
Ledig Female

Omega

Hindi
Ledig Male

Sara

Italian
Ledig Female

Sarah

English
Ledig Female

Siwis

French
Ledig Female

Sky

English
Ledig Female

Xiaobei

Chinese
Ledig Female

Xiaoni

Chinese
Ledig Female

Xiaoxiao

Chinese
Ledig Female

Yunjian

Chinese
Ledig Male

Kokoro TTS — OSS

Yes. Kokoro is released under the Apache 2.0 license and sits in the free tier, so it can be used commercially at no cost.

Kokoro is only 82M parameters and uses an efficient StyleTTS + ISTFTNet design, letting it generate audio nearly 100x faster than real-time on a GPU with just ~1.5GB VRAM while staying natural-sounding.

Kokoro supports English, Japanese, Chinese, French, Italian, Portuguese, Spanish, and Hindi, each with its own expressive voicepacks. It does not support voice cloning.
← Alle stemmer