Kitten TTS TTS
An ultra-lightweight ONNX model under 80MB that runs high-quality TTS on CPU with no GPU at all.
Whāriki i tōna kupu i roto i ngā tohu SSML mō te whakahaere tika:
<speak><prosody rate="slow">Slow speech</prosody></speak>
E mōhio ana ngā tohu ki te tauira i kōwhiria - ka kōwhiria kia whakawātea tētahi ki roto i tōna kupu i reira ka puta ai:
Ka pānui tēnei tauira i te kupu noa, nā reira ka whakakāhoretia ngā tohu ā-waitara. Mō te āhua o te tohu-taihi, ka huri ki tētahi tauira whakamārama pēnei i a Orpheus, Bark rānei.
Ka tautuhia ngā tohutohu ā-ringa (wāhi = tohutohu):
Mo Kitten TTS
Kitten TTS by KittenML is built for extreme portability: an ONNX-based model with variants from 15M to 80M parameters, weighing just 25-80 MB on disk. It runs entirely on CPU with 0 VRAM, yet ships 8 built-in voices, adjustable speech speed, and built-in text preprocessing that handles numbers, currencies, and units before synthesis. Output is 24kHz. While its quality sits below the larger models, its tiny size and fast (~2s) CPU inference make it ideal for edge deployment and latency-sensitive applications that can't assume a GPU is present.
Pai mo: Fast lightweight TTS, edge deployment, low-latency applications
Ka tirohia katoa Kitten TTS ngā oroI te tirohanga
- Ka whakawhanakehia
- KittenML
- Ka taea te whakawātea
- Apache 2.0
- Karaka
- free
- Āhuatanga
- fast
- Whakakōrero reo
- Kāore
- reo
- English
- Kāri nui rawa
- 5000