Chinese (Mandarin) Umbhalo ukuya kuSpeech
Ujikelezo Chinese (Mandarin) amagama aqhelekileyo ngeelizwi ze-AI. 25 iilizwi. Isimahla, akukho ubhaliso — khuphela njenge MP3 okanye WAV.
Ulawulo oluchanekileyo:
<speak><prosody rate="slow">Slow speech</prosody></speak>
Ii-tags imodeli ekhethiweyo iqonda - nqakraza ukushiya enye kumbhalo wakho apho isenza khona:
Le modeli ifunda umbhalo oqhelekileyo, ngoko ke i-inline tags ilahleka. Uphawu olusekelwe kwi-emotions, tshintshela kwimodeli ebonisa umbono njenge-Orpheus okanye i-Bark.
Chaza ubeko lwephepha
I-About Chinese (Mandarin) Umbhalo ukuya kuthetha
Mandarin text-to-speech lives or dies on tone: it has four lexical tones plus a neutral tone, and getting the contour wrong turns "mā" (mother) into "mǎ" (horse), so the model must predict pitch per syllable, not just per sentence. Tone sandhi adds another layer — for example two third tones in a row shift the first to a rising tone, and the words "一" (yī) and "不" (bù) change tone depending on what follows. Because Hanzi carry no spaces and many characters are polyphonic (多音字), high-quality Chinese synthesis depends heavily on word segmentation and grapheme-to-phoneme disambiguation from context.
Iinketho ze projekti — 中文(普通话)
“今天天气很好,我们一起去公园散步,顺便买点水果回家吧。”
- Igama eliqhelekileyo
- 中文(普通话)
- Abathethi
- about 1.1 billion speakers (roughly 920 million native Mandarin)
- Usapho lwesiNgesi
- Sinitic branch of Sino-Tibetan
- Igama lefayile le CVS:
- Chinese characters (Hanzi) — Simplified and Traditional
- Ithetha
- Mainland China, Taiwan, Singapore, Malaysia, Hong Kong, global Chinese diaspora