VieNeu-TTS-v2 टीटीएस
A Vietnamese-first, CPU-only model with en-vi code-switching, 7 regional preset voices, and zero-shot cloning.
अचूक नियंत्रण करीता SSML टॅग अंतर्गत पाठ्य वेल्ड करा:
<speak><prosody rate="slow">Slow speech</prosody></speak>
निवडलेले नमूना समजून घेणारे टॅग - पाठ्य अंतर्गत एक टाका जेथे ते घडते:
हे मॉडेल सादा पाठ्य वाचते, त्यामुळे इनलाईन टॅग दुर्लक्ष केले जातात. टॅग आधारीत भावना करीता, Orpheus किंवा Bark सारखे अभिव्यक्ती मॉडेल करीता बदलवा.
इच्छिक उच्चारण निश्चित करा (शब्द = उच्चारण):
विषयी VieNeu-TTS-v2
VieNeu-TTS-v2 is a 300M-parameter Vietnamese-first model built on a Qwen3 backbone and trained on more than 10,000 hours of bilingual data. It handles seamless English-Vietnamese code-switching, ships 7 preset voices spanning Northern and Southern accents, and clones a voice instantly from just 3-5 seconds of reference audio. Notably it runs entirely on CPU — using GGUF Q4 inference plus an ONNX audio decoder — with no GPU required, finishing a generation in about 7 seconds. It's purpose-built for Vietnamese content and bilingual en-vi narration, an underserved niche in open TTS.
सर्वोत्तम: Vietnamese content and bilingual en-vi narration
सर्व ब्राऊज करा VieNeu-TTS-v2 आवाजएक नजर
- डेव्हलपर
- Phạm Nguyễn Ngọc Bảo
- परवाना
- Apache 2.0
- टर
- standard
- वेग
- fast
- आवाज क्लोन
- होय
- भाषाName
- Vietnamese, English
- कमाल अक्षरे
- 1000