Chinese (Mandarin) Àkọlé sí Àkọ́kọ́
Ìjánú àtòjọ-ẹ̀yàn Chinese (Mandarin) kọ́ àkọlé sí àkọlé àìdálẹ̀ láti inú àwọn ìrànwọ́ AI. 25 Àwọn àwòrán. Ko si ifẹ, ko si iforukọsilẹ — gba pada bi MP3 tabi WAV.
Fi àkọlé rẹ pamọ́ sí àwọn àmì-ìwé SSML fún ìdáràn:
<speak><prosody rate="slow">Slow speech</prosody></speak>
Àwọn Àmì-ìwé tí àwọn ìṣàmúlò-ètò tí a yàn gbọ́ - tẹ̀ láti fi ọkan sínú àkọ́lé rẹ̀ nínú àwọn ààyè-iṣẹ́ tí o bá jẹ́:
Àwọn àwọn àkọlé àwọn ààyè-iṣẹ́ àwọn àwọn àmì-ìwé àwọn àmì-ìwé àwọn àwọn àmì-ìwé àwọn à
Àwọn àwọn ìṣàfarawé àwọn àwọn ìṣàfarawé àwọn (ọrọ = ìṣàfàlì):
Ààyè-iṣẹ́ Chinese (Mandarin) Àkọ́lé sí àkọ́lé
Mandarin text-to-speech lives or dies on tone: it has four lexical tones plus a neutral tone, and getting the contour wrong turns "mā" (mother) into "mǎ" (horse), so the model must predict pitch per syllable, not just per sentence. Tone sandhi adds another layer — for example two third tones in a row shift the first to a rising tone, and the words "一" (yī) and "不" (bù) change tone depending on what follows. Because Hanzi carry no spaces and many characters are polyphonic (多音字), high-quality Chinese synthesis depends heavily on word segmentation and grapheme-to-phoneme disambiguation from context.
Àwọn Ààyè-iṣẹ́ — 中文(普通话)
“今天天气很好,我们一起去公园散步,顺便买点水果回家吧。”
- Orúkọ̀
- 中文(普通话)
- Àwọn Àkọlé
- about 1.1 billion speakers (roughly 920 million native Mandarin)
- Àwọn
- Sinitic branch of Sino-Tibetan
- Àwọn Àkọlé
- Chinese characters (Hanzi) — Simplified and Traditional
- Tí a Fẹ̀
- Mainland China, Taiwan, Singapore, Malaysia, Hong Kong, global Chinese diaspora