VieNeu-TTS-v2 TTS
A Vietnamese-first, CPU-only model with en-vi code-switching, 7 regional preset voices, and zero-shot cloning.
Wrap wanu malemba mu SSML tags kwa kuwongolera moyenera:
<speak><prosody rate="slow">Slow speech</prosody></speak>
Tags chosankhidwa chitsanzo amamvetsa - dinani kuti aphe mmodzi m'mawu anu pamene chimachitika:
Izi ndi njira yolemba malemba oyera, kotero ma tag ophatikizidwa amasiya kuganiziridwa. Kuti mupange ma tag ogwirizana ndi maganizo, gwiritsani ntchito njira yolemba malemba monga Orpheus kapena Bark.
Define custom pronunciations (word = pronunciation):
Za VieNeu-TTS-v2
VieNeu-TTS-v2 is a 300M-parameter Vietnamese-first model built on a Qwen3 backbone and trained on more than 10,000 hours of bilingual data. It handles seamless English-Vietnamese code-switching, ships 7 preset voices spanning Northern and Southern accents, and clones a voice instantly from just 3-5 seconds of reference audio. Notably it runs entirely on CPU — using GGUF Q4 inference plus an ONNX audio decoder — with no GPU required, finishing a generation in about 7 seconds. It's purpose-built for Vietnamese content and bilingual en-vi narration, an underserved niche in open TTS.
Best kwa: Vietnamese content and bilingual en-vi narration
Pezani zonse VieNeu-TTS-v2 maganizoPa mphindi
- Wopanga
- Phạm Nguyễn Ngọc Bảo
- License
- Apache 2.0
- Mtundu
- standard
- Kuyenda
- fast
- Kusintha kwa mawu
- Yes
- Zilankhulo
- Vietnamese, English
- Max characters
- 1000