VieNeu-TTS-v2

VieNeu-TTS-v2 音声翻訳

A Vietnamese-first, CPU-only model with en-vi code-switching, 7 regional preset voices, and zero-shot cloning.

登録 5000文字の制限を設けました

SSML タグでテキストを囲み、正確な制御を行う:

<speak><prosody rate="slow">Slow speech</prosody></speak>

選択したモデルが理解するタグ - クリックしてテキストにドラッグします:

このモデルは単純テキストを読み込み、インラインタグは無視されます。タグベースの感情を表現するには、Orpheus や Bark のような表現モデルに切り替えてください。

カスタム発音を定義 (単語=発音):

-12 +12
0.5x 2.0x
ピパー、VITS、MeloTTS をフリーで使用
生成したオーディオがここに表示されます。モデルを選択し、テキストを入力して、生成をクリックします。
オーディオを作成しましたName
0:00
音声をダウンロード ダウンロード リンクは24時間で失効します
無料階級:個人用。 商用ライセンス $5/月から
これを自分の声にしよう 30秒で声をクローン
TTS.aiが気に入りましたか?友達に教えてあげましょう!

情報 VieNeu-TTS-v2

VieNeu-TTS-v2 is a 300M-parameter Vietnamese-first model built on a Qwen3 backbone and trained on more than 10,000 hours of bilingual data. It handles seamless English-Vietnamese code-switching, ships 7 preset voices spanning Northern and Southern accents, and clones a voice instantly from just 3-5 seconds of reference audio. Notably it runs entirely on CPU — using GGUF Q4 inference plus an ONNX audio decoder — with no GPU required, finishing a generation in about 7 seconds. It's purpose-built for Vietnamese content and bilingual en-vi narration, an underserved niche in open TTS.

適合する: Vietnamese content and bilingual en-vi narration

すべてブラウズ VieNeu-TTS-v2 声

概要

開発者
Phạm Nguyễn Ngọc Bảo
ライセンス
Apache 2.0
動物
standard
スピード
fast
声のクローン
はい
言語
Vietnamese, English
最大文字数
1000

VieNeu-TTS-v2 声

Bích Ngọc (North, Female)

Vietnamese
標準 Female

Phạm Tuyên (North, Male)

Vietnamese
標準 Male

Thanh Bình (North, Male)

Vietnamese
標準 Male

Thái Sơn (South, Male)

Vietnamese
標準 Male

Thục Đoan (South, Female)

Vietnamese
標準 Female

Trúc Ly (North, Female)

Vietnamese
標準 Female

Xuân Vĩnh (South, Male)

Vietnamese
標準 Male

VieNeu-TTS-v2 よくある質問

Yes. VieNeu-TTS-v2 runs entirely on CPU via GGUF Q4 inference and an ONNX audio decoder — no GPU needed — and completes a generation in around 7 seconds.

It is Vietnamese-first with English support and seamless en-vi code-switching. It ships 7 preset voices spanning Northern and Southern Vietnamese accents.

Yes. It supports instant zero-shot voice cloning from just 3-5 seconds of reference audio. It is Apache 2.0 licensed and free to use commercially.
← すべての声