Chinese (Mandarin) テキストから音声へ

ターン Chinese (Mandarin) テキストから自然な音声に変換することができます 25 声. 無料で登録なし — MP3 または WAV でダウンロード。

登録 5000文字の制限を設けました

SSML タグでテキストを囲み、正確な制御を行う:

<speak><prosody rate="slow">Slow speech</prosody></speak>

選択したモデルが理解するタグ - クリックしてテキストにドラッグします:

このモデルは単純テキストを読み込み、インラインタグは無視されます。タグベースの感情を表現するには、Orpheus や Bark のような表現モデルに切り替えてください。

カスタム発音を定義 (単語=発音):

-12 +12
0.5x 2.0x
ピパー、VITS、MeloTTS をフリーで使用
生成したオーディオがここに表示されます。モデルを選択し、テキストを入力して、生成をクリックします。
オーディオを作成しましたName
0:00
音声をダウンロード ダウンロード リンクは24時間で失効します
無料階級:個人用。 商用ライセンス $5/月から
これを自分の声にしよう 30秒で声をクローン
TTS.aiが気に入りましたか?友達に教えてあげましょう!

情報 Chinese (Mandarin) テキストから音声を生成する

Mandarin text-to-speech lives or dies on tone: it has four lexical tones plus a neutral tone, and getting the contour wrong turns "mā" (mother) into "mǎ" (horse), so the model must predict pitch per syllable, not just per sentence. Tone sandhi adds another layer — for example two third tones in a row shift the first to a rising tone, and the words "一" (yī) and "不" (bù) change tone depending on what follows. Because Hanzi carry no spaces and many characters are polyphonic (多音字), high-quality Chinese synthesis depends heavily on word segmentation and grapheme-to-phoneme disambiguation from context.

サンプル — 中文(普通话)

“今天天气很好,我们一起去公园散步,顺便买点水果回家吧。”

実名
中文(普通话)
スピーカー
about 1.1 billion speakers (roughly 920 million native Mandarin)
言語族
Sinitic branch of Sino-Tibetan
スクリプト
Chinese characters (Hanzi) — Simplified and Traditional
話した言葉
Mainland China, Taiwan, Singapore, Malaysia, Hong Kong, global Chinese diaspora

25 Chinese (Mandarin) 声

Chinese Speaker 1

Bark
標準 Neutral

Chinese Speaker 2

Bark
標準 Neutral

Chinese Speaker

Bark Small
標準 Neutral

Chinese Female

CosyVoice 2
標準 Female

Chinese Male

CosyVoice 2
標準 Male

Chinese Female

CosyVoice3
標準 Female

Chinese Male

CosyVoice3
標準 Male

Default (Chinese)

Darwin TTS
標準 Neutral

Default

GPT-SoVITS
標準 Neutral

Chinese Default

IndexTTS-2
標準 Neutral

Xiaobei

Kokoro
自由 Female

Xiaoni

Kokoro
自由 Female

Xiaoxiao

Kokoro
自由 Female

Yunjian

Kokoro
自由 Male

Chinese

MeloTTS
自由 Female

Default (Chinese)

Ming-Omni TTS
自由 Neutral

Chinese

MOSS-TTS Nano
標準 Neutral

Default (Chinese)

MOSS-TTSD
標準 Neutral

Chinese

OpenVoice
プレミアム Neutral

Huayan (Chinese)

Piper
自由 Female

Uncle Fu

Qwen3 TTS
標準 Male

Chinese Default

Spark TTS
標準 Neutral

Speaker 1 (Chinese)

VibeVoice
標準 Neutral

Speaker 2 (Chinese)

VibeVoice
標準 Neutral

Default Chinese

VoxCPM
標準 Neutral

人々が使うもの Chinese (Mandarin) テキストから音声を生成する

E-learning and Mandarin language-teaching narration
Short-video (Douyin/Bilibili) and livestream voiceover
Navigation and in-car voice prompts
Customer-service IVR and chatbot voices
News and audiobook narration for the diaspora

Chinese (Mandarin) テキストから音声へ

Yes. You can paste either Simplified (mainland/Singapore) or Traditional (Taiwan/Hong Kong) text; both are read in Mandarin pronunciation.

The model predicts each syllable's tone contour from context and applies tone sandhi rules, so sequences like third-tone pairs and the special cases of 一 and 不 come out naturally.

Mostly yes. Multi-reading characters such as 行 (xíng vs háng) or 长 (cháng vs zhǎng) are disambiguated from surrounding words, though rare proper nouns can still be ambiguous.

These voices are Standard Mandarin (Putonghua). Cantonese uses a different tone system and pronunciation and is not the same as Mandarin TTS.

関連言語