VieNeu-TTS-v2 TTS
A Vietnamese-first, CPU-only model with en-vi code-switching, 7 regional preset voices, and zero-shot cloning.
Wrap ou tèks nan SSML tags pou presizyon kontwòl:
<speak><prosody rate="slow">Slow speech</prosody></speak>
Tags ke modèl la chwazi konprann — klike pou mete yon nan tèks ou kote li rive:
Modèl sa a li tèks senp, se poutèt sa atik ki nan liy yo pa pran an kont. Pou efè ki baze sou atik, chanje pou yon modèl ekspresyon tankou Orpheus oswa Bark.
Define prononciations Custom (mot = prononciation):
Atik VieNeu-TTS-v2
VieNeu-TTS-v2 is a 300M-parameter Vietnamese-first model built on a Qwen3 backbone and trained on more than 10,000 hours of bilingual data. It handles seamless English-Vietnamese code-switching, ships 7 preset voices spanning Northern and Southern accents, and clones a voice instantly from just 3-5 seconds of reference audio. Notably it runs entirely on CPU — using GGUF Q4 inference plus an ONNX audio decoder — with no GPU required, finishing a generation in about 7 seconds. It's purpose-built for Vietnamese content and bilingual en-vi narration, an underserved niche in open TTS.
Pi bon pou: Vietnamese content and bilingual en-vi narration
Navigue tout VieNeu-TTS-v2 VoyYon ti gade
- Pwogramè
- Phạm Nguyễn Ngọc Bảo
- Lisans
- Apache 2.0
- Nivo
- standard
- Vitès
- fast
- Klonaj vwa
- Wi
- Lang
- Vietnamese, English
- Karakteris maksimòm
- 1000