Chinese (Mandarin) Text in die Rede

Drehen Chinese (Mandarin) Text in natürliche Sprache mit KI-Stimmen. 25 Stimmen. Kostenlos, ohne Anmeldung – als MP3 oder WAV herunterladen.

Melden Sie sich an für 5.000 Zeichen-Grenze

Verpacken Sie Ihren Text in SSML-Tags für eine präzise Kontrolle:

<speak><prosody rate="slow">Slow speech</prosody></speak>

Tags, die das ausgewählte Modell versteht — klicken Sie, um einen in Ihren Text zu legen, wo es passiert:

Dieses Modell liest Text, so dass Inline-Tags ignoriert werden. Für tag-basierte Emotion, wechseln Sie zu einem ausdrucksstarken Modell wie Orpheus oder Bark.

Benutzerdefinierte Aussprachen definieren (Wort = Aussprache):

-12 +12
0.5x 2.0x
Frei mit Piper, VITS, MeloTTS
Hier erscheint Ihr generiertes Audio. Wählen Sie ein Modell, geben Sie Text ein und klicken Sie auf Generieren.
Audio-Erzeugung erfolgreich
0:00
Audio herunterladen Download.srt Link läuft in 24h aus
Freier Dienstgrad: persönlicher Gebrauch. Kommerzielle Lizenz ab $5/mo
Gefällt dir TTS.ai? Erzähl es deinen Freunden!

Über Chinese (Mandarin) Text in die Rede

Mandarin text-to-speech lives or dies on tone: it has four lexical tones plus a neutral tone, and getting the contour wrong turns "mā" (mother) into "mǎ" (horse), so the model must predict pitch per syllable, not just per sentence. Tone sandhi adds another layer — for example two third tones in a row shift the first to a rising tone, and the words "一" (yī) and "不" (bù) change tone depending on what follows. Because Hanzi carry no spaces and many characters are polyphonic (多音字), high-quality Chinese synthesis depends heavily on word segmentation and grapheme-to-phoneme disambiguation from context.

Stichprobe — 中文(普通话)

“今天天气很好,我们一起去公园散步,顺便买点水果回家吧。”

Einheimischer Name
中文(普通话)
Redner
about 1.1 billion speakers (roughly 920 million native Mandarin)
Sprachfamilie
Sinitic branch of Sino-Tibetan
Skript
Chinese characters (Hanzi) — Simplified and Traditional
Gesprochen
Mainland China, Taiwan, Singapore, Malaysia, Hong Kong, global Chinese diaspora

25 Chinese (Mandarin) Stimmen

Chinese Speaker 1

Bark
Standard Neutral

Chinese Speaker 2

Bark
Standard Neutral

Chinese Speaker

Bark Small
Standard Neutral

Chinese Female

CosyVoice 2
Standard Female

Chinese Male

CosyVoice 2
Standard Male

Chinese Female

CosyVoice3
Standard Female

Chinese Male

CosyVoice3
Standard Male

Default (Chinese)

Darwin TTS
Standard Neutral

Default

GPT-SoVITS
Standard Neutral

Chinese Default

IndexTTS-2
Standard Neutral

Xiaobei

Kokoro
Frei Female

Xiaoni

Kokoro
Frei Female

Xiaoxiao

Kokoro
Frei Female

Yunjian

Kokoro
Frei Male

Chinese

MeloTTS
Frei Female

Default (Chinese)

Ming-Omni TTS
Frei Neutral

Chinese

MOSS-TTS Nano
Standard Neutral

Default (Chinese)

MOSS-TTSD
Standard Neutral

Chinese

OpenVoice
Prämie Neutral

Huayan (Chinese)

Piper
Frei Female

Uncle Fu

Qwen3 TTS
Standard Male

Chinese Default

Spark TTS
Standard Neutral

Speaker 1 (Chinese)

VibeVoice
Standard Neutral

Speaker 2 (Chinese)

VibeVoice
Standard Neutral

Default Chinese

VoxCPM
Standard Neutral

Was die Menschen benutzen Chinese (Mandarin) Text zu Rede für

E-learning and Mandarin language-teaching narration
Short-video (Douyin/Bilibili) and livestream voiceover
Navigation and in-car voice prompts
Customer-service IVR and chatbot voices
News and audiobook narration for the diaspora

Chinese (Mandarin) Text in die Rede — FAQ

Yes. You can paste either Simplified (mainland/Singapore) or Traditional (Taiwan/Hong Kong) text; both are read in Mandarin pronunciation.

The model predicts each syllable's tone contour from context and applies tone sandhi rules, so sequences like third-tone pairs and the special cases of 一 and 不 come out naturally.

Mostly yes. Multi-reading characters such as 行 (xíng vs háng) or 长 (cháng vs zhǎng) are disambiguated from surrounding words, though rare proper nouns can still be ambiguous.

These voices are Standard Mandarin (Putonghua). Cantonese uses a different tone system and pronunciation and is not the same as Mandarin TTS.

Verwandte Sprachen