Kokoro

Kokoro TTS

An 82M-parameter open model from Hexgrad that delivers studio-quality speech at nearly 100x real-time.

Signa per 5000 caràcters límit

Ajusta el text a les etiquetes SSML per al control precís:

<speak><prosody rate="slow">Slow speech</prosody></speak>

Etiquetes del model seleccionat entenen el clic show clic per a deixar- ne un al text a on succeeix:

Aquest model llegeix text pla, així que les etiquetes inserides s' ignoren. Per a emocions basades en etiquetes, canvieu a un model expressiu com Orfeus o Bark.

Defineix pronúncies personalitzades (word = pronunciació):

-12 +12
0.5x 2.0x
Lliure amb Pipista, VITS, MeloTTS
Aquí apareixerà el vostre àudio generat. Escolliu un model, introduïu text i cliqueu Genera.
L' àudio s' ha generat correctament
0:00
Descarrega àudio Descarrega.srt L' enllaç expirarà el 24h
Correlitzador lliure: ús personal. Llicència de venda de 5/mo
Fes que això sigui la teva pròpia veu Clona una veu en 30 segons
Els teus amics!

Quant a Kokoro

Kokoro, built by Hexgrad, is a deliberately tiny 82-million-parameter model that punches far above its size class. It uses a StyleTTS + ISTFTNet architecture and was trained on roughly 1,200 hours of speech, yet generates audio close to 100x faster than real-time on a GPU while staying remarkably natural and expressive. It covers English, Japanese, Chinese, French, Italian, Portuguese, Spanish, and Hindi with a varied set of expressive voicepacks, and supports streaming. The combination of small footprint, low latency, and a permissive license has made Kokoro one of the most popular free models for high-volume and streaming use — it carries the largest share of traffic on TTS.ai.

Millor per: High-quality TTS with minimal latency, streaming applications

Navega- ho tot Kokoro veus

En una mirada

Desenvolupador
Hexgrad
Llicència
Apache 2.0
TierCity name (optional, probably does not need a translation)
free
Velocitat
fast
clonació de veu
No
Idiomes
English, Japanese, Chinese, French, Italian, Portuguese, Spanish, Hindi
Nombre màxim de caràcters
500

Kokoro veus

Adam

English
Lliure Male

Alex

Spanish
Lliure Male

Alex

Portuguese
Lliure Male

Alpha

Japanese
Lliure Female

Alpha

Hindi
Lliure Female

Bella

English
Lliure Female

Dora

Portuguese
Lliure Female

Dora

Spanish
Lliure Female

Emma (British)

English
Lliure Female

George (British)

English
Lliure Male

Gongitsune

Japanese
Lliure Female

Heart

English
Lliure Female

Isabella (British)

English
Lliure Female

Lewis (British)

English
Lliure Male

Michael

English
Lliure Male

Nicola

Italian
Lliure Male

Nicole

English
Lliure Female

Omega

Hindi
Lliure Male

Sara

Italian
Lliure Female

Sarah

English
Lliure Female

Siwis

French
Lliure Female

Sky

English
Lliure Female

Xiaobei

Chinese
Lliure Female

Xiaoni

Chinese
Lliure Female

Xiaoxiao

Chinese
Lliure Female

Yunjian

Chinese
Lliure Male

Kokoro PMF TTS

Yes. Kokoro is released under the Apache 2.0 license and sits in the free tier, so it can be used commercially at no cost.

Kokoro is only 82M parameters and uses an efficient StyleTTS + ISTFTNet design, letting it generate audio nearly 100x faster than real-time on a GPU with just ~1.5GB VRAM while staying natural-sounding.

Kokoro supports English, Japanese, Chinese, French, Italian, Portuguese, Spanish, and Hindi, each with its own expressive voicepacks. It does not support voice cloning.
← Totes les veus