VieNeu-TTS-v2

VieNeu-TTS-v2 TTS

A Vietnamese-first, CPU-only model with en-vi code-switching, 7 regional preset voices, and zero-shot cloning.

ምዝገባ 5,000 ካራቴግራም

SSML tags ሒዝካ ጽሑፍካ ሒዝካ ንምውሳድ:

<speak><prosody rate="slow">Slow speech</prosody></speak>

ርኢቶታት እቲ ዝተመርጸ ሞዴል ዝፈልጦ — ጠቅልል ንኸውዕሎ ኣብ ጽሑፍካ ኣብ ዝግበር ቦታ:

እዚ ሞዴል'ዚ ጽሑፍ ቀሊል ይንብብ፣ ከምኡ'ውን ኣብ መስመር ዝርከብ ቴግታት ይቕረ ይበሃል። ን tag-based emotion፣ ናብ ሞዴል ስነ-ኣእምሮኣዊ ከም Orpheus ወይ Bark ምቕያር

ድምጺ ተለፎን (ቓል = ድምጺ)

-12 +12
0.5x 2.0x
ነጻ ምስ Piper, VITS, MeloTTS
እቲ ዝተፈጠረ ድምጺ እዚኣ ይርከብ። ሓደ ሞዴል ምረጽ፣ ጽሑፍ ኣትሒዝካ፣ ንተፈጥሮ ጠቅልል።
ድምጺ ተኸፊሉ
0:00
ድምጺ መዝጊብ .srt መዝገብ 24 ሰዓታት
ነጻ ደረጃ: ንባዕልኻ ንምጥቃም. ውልቀ-መዚ ውልቀ-መዚ ካብ $5/month
TTS.ai ትወዳደሮ? ንፍቓድካ ንገረኒ!

ብዛዕባ VieNeu-TTS-v2

VieNeu-TTS-v2 is a 300M-parameter Vietnamese-first model built on a Qwen3 backbone and trained on more than 10,000 hours of bilingual data. It handles seamless English-Vietnamese code-switching, ships 7 preset voices spanning Northern and Southern accents, and clones a voice instantly from just 3-5 seconds of reference audio. Notably it runs entirely on CPU — using GGUF Q4 inference plus an ONNX audio decoder — with no GPU required, finishing a generation in about 7 seconds. It's purpose-built for Vietnamese content and bilingual en-vi narration, an underserved niche in open TTS.

ምርኣይ: Vietnamese content and bilingual en-vi narration

ርአ VieNeu-TTS-v2 ቃላት

ኣብ ሓደ ገጽ

መተግበሪያ
Phạm Nguyễn Ngọc Bảo
ውልቀ-መዚ
Apache 2.0
ቍጽሪ
standard
ፍጥነት
fast
ድምጺ
ኣየ
ቋንቋ
Vietnamese, English
ቍጽሪ ኣርእስቲ
1000

VieNeu-TTS-v2 ቃላት

Bích Ngọc (North, Female)

Vietnamese
ቍጽሪ Female

Phạm Tuyên (North, Male)

Vietnamese
ቍጽሪ Male

Thanh Bình (North, Male)

Vietnamese
ቍጽሪ Male

Thái Sơn (South, Male)

Vietnamese
ቍጽሪ Male

Thục Đoan (South, Female)

Vietnamese
ቍጽሪ Female

Trúc Ly (North, Female)

Vietnamese
ቍጽሪ Female

Xuân Vĩnh (South, Male)

Vietnamese
ቍጽሪ Male

VieNeu-TTS-v2 ሕቶታት ዝንበረሉ

Yes. VieNeu-TTS-v2 runs entirely on CPU via GGUF Q4 inference and an ONNX audio decoder — no GPU needed — and completes a generation in around 7 seconds.

It is Vietnamese-first with English support and seamless en-vi code-switching. It ships 7 preset voices spanning Northern and Southern Vietnamese accents.

Yes. It supports instant zero-shot voice cloning from just 3-5 seconds of reference audio. It is Apache 2.0 licensed and free to use commercially.
← ድምጺ