Kani TTS 2

Kani TTS 2 TTS

An ultra-lightweight 400M English model that runs in just 3GB of VRAM at a 0.2 real-time factor.

Înscrie-te pentru limitele de 5000 de caractere

Întoarceți textul în etichetele SSML pentru un control precis:

<speak><prosody rate="slow">Slow speech</prosody></speak>

Etichetele modelului selectat înțeleg — click pentru a lăsa unul în textul tău unde se întâmplă:

Acest model citește textul simplu, astfel încât etichetele inline sunt ignorate. Pentru emoții bazate pe tag, schimbați la un model expresiv cum ar fi Orpheus sau Bark.

Definiți pronunțiare personalizată (cuvânt = pronunție):

-12 +12
0.5x 2.0x
Gratuit cu Piper, VITS, MeloTTS
Audio generat va apărea aici. Alegeți un model, introduceți text și faceți clic pe Generați.
Audio generat cu succes
0:00
Descarcă audio Descărcare.srt Legătura expiră în 24 ore
Gratuit: utilizare personală. Licență comercială de la 5$/mo
Spune-i prietenilor tăi!

Despre Kani TTS 2

Kani-TTS-2 by NineNineSix is an ultra-lightweight 400M-parameter text-to-speech model built on a Liquid AI LFM2 backbone with NVIDIA's NanoCodec. It runs in just 3GB of VRAM and generates roughly ten seconds of speech in about two seconds on an A100 — a real-time factor near 0.2. The current public release ships an English-only checkpoint and, unlike its predecessor, does not expose the speaker-embedding hook needed for voice cloning. Its strength is fast, low-cost English generation on modest hardware, which makes it a good fit for quick previews and high-volume English narration. It is released under Apache 2.0 and offered on the free tier.

Cel mai bun pentru: Fast English generation on low-VRAM hardware, quick previews

Navigați toate Kani TTS 2 voci

La o privire

Dezvoltator
NineNineSix
Licență
Apache 2.0
Nivel
free
Viteză
fast
Clonarea vocală
Nu.
Limbi
English
Caractere maxime
1000

Kani TTS 2 voci

Default

English
Standard Neutral

Kani TTS 2 TTS – FAQ

It runs in just 3GB of VRAM and produces about ten seconds of speech in roughly two seconds on an A100 — a real-time factor near 0.2 — thanks to its 400M-parameter LFM2 backbone and NanoCodec.

No. The current v2 release removed the public speaker-embedding hook, so cloning is not available. For cloning, use Chatterbox, IndexTTS-2, or GPT-SoVITS.

English only. The public release ships a single English checkpoint; for non-English speech, use a model like Kokoro or MeloTTS.
← Toate vocile