Chinese (Mandarin) 文本到语音

转向 Chinese (Mandarin) 以 AI 声音 进入自然语言的文本 。 25 声音. 免费,没有注册——下载为 MP3 或 WAV 。

签名 对 5,000 字符限制的 5 000 个字符

在 SSML 标记中折行文本以精确控制 :

<speak><prosody rate="slow">Slow speech</prosody></speak>

标记选中模式的理解度 - 单击将一个输入到文本中, 发生时 :

这个模型读的是简单的文字, 所以内嵌标签会被忽略。 对于基于标签的情感, 请切换到像 Orpheus 或 Bark 这样的表达模式 。

定义自定义发音( Word = 发音) :

-12 +12
0.5x 2.0x
免费的管道、VITS、MelotTS
您生成的音频将在此显示。 选择一个模型, 输入文本, 并单击生成 。
音频生成成功
0:00
下载音频 下载.strt 24小时后链接过期
免费:个人使用。 5美元/美元商业许可证
使这个声音成为你自己的声音 30秒后打开声音
喜欢TTS.ai吗?告诉你的朋友吧!

关于 Chinese (Mandarin) 文本到语音

Mandarin text-to-speech lives or dies on tone: it has four lexical tones plus a neutral tone, and getting the contour wrong turns "mā" (mother) into "mǎ" (horse), so the model must predict pitch per syllable, not just per sentence. Tone sandhi adds another layer — for example two third tones in a row shift the first to a rising tone, and the words "一" (yī) and "不" (bù) change tone depending on what follows. Because Hanzi carry no spaces and many characters are polyphonic (多音字), high-quality Chinese synthesis depends heavily on word segmentation and grapheme-to-phoneme disambiguation from context.

抽样样本 — 中文(普通话)

“今天天气很好,我们一起去公园散步,顺便买点水果回家吧。”

原土著名称
中文(普通话)
发言者
about 1.1 billion speakers (roughly 920 million native Mandarin)
语言家庭
Sinitic branch of Sino-Tibetan
脚本
Chinese characters (Hanzi) — Simplified and Traditional
说到
Mainland China, Taiwan, Singapore, Malaysia, Hong Kong, global Chinese diaspora

25 Chinese (Mandarin) 声音

Chinese Speaker 1

Bark
标准 Neutral

Chinese Speaker 2

Bark
标准 Neutral

Chinese Speaker

Bark Small
标准 Neutral

Chinese Female

CosyVoice 2
标准 Female

Chinese Male

CosyVoice 2
标准 Male

Chinese Female

CosyVoice3
标准 Female

Chinese Male

CosyVoice3
标准 Male

Default (Chinese)

Darwin TTS
标准 Neutral

Default

GPT-SoVITS
标准 Neutral

Chinese Default

IndexTTS-2
标准 Neutral

Xiaobei

Kokoro
自由 Female

Xiaoni

Kokoro
自由 Female

Xiaoxiao

Kokoro
自由 Female

Yunjian

Kokoro
自由 Male

Chinese

MeloTTS
自由 Female

Default (Chinese)

Ming-Omni TTS
自由 Neutral

Chinese

MOSS-TTS Nano
标准 Neutral

Default (Chinese)

MOSS-TTSD
标准 Neutral

Chinese

OpenVoice
[Translation temporarily unavailable. Please try again.] Neutral

Huayan (Chinese)

Piper
自由 Female

Uncle Fu

Qwen3 TTS
标准 Male

Chinese Default

Spark TTS
标准 Neutral

Speaker 1 (Chinese)

VibeVoice
标准 Neutral

Speaker 2 (Chinese)

VibeVoice
标准 Neutral

Default Chinese

VoxCPM
标准 Neutral

人使用什么 Chinese (Mandarin) 文本到 语音的文本

E-learning and Mandarin language-teaching narration
Short-video (Douyin/Bilibili) and livestream voiceover
Navigation and in-car voice prompts
Customer-service IVR and chatbot voices
News and audiobook narration for the diaspora

Chinese (Mandarin) 文本到语音 - FAQ

Yes. You can paste either Simplified (mainland/Singapore) or Traditional (Taiwan/Hong Kong) text; both are read in Mandarin pronunciation.

The model predicts each syllable's tone contour from context and applies tone sandhi rules, so sequences like third-tone pairs and the special cases of 一 and 不 come out naturally.

Mostly yes. Multi-reading characters such as 行 (xíng vs háng) or 长 (cháng vs zhǎng) are disambiguated from surrounding words, though rare proper nouns can still be ambiguous.

These voices are Standard Mandarin (Putonghua). Cantonese uses a different tone system and pronunciation and is not the same as Mandarin TTS.

相关语言