KhanomTan TTS TTS
An open, commercially-licensed Thai-first text-to-speech model with multiple speaker voices.
නිවැරදි පාලනය සඳහා SSML ටැග් ඔබේ පෙළ ආවරණය:
<speak><prosody rate="slow">Slow speech</prosody></speak>
තෝරාගත් ආකෘතිය තේරුම් ටැග් - එය සිදුවන තැන ඔබේ පෙළ එක් වැටීමට ක්ලික් කරන්න:
මෙම ආකෘතිය සරල පෙළ කියවනවා, ඒ නිසා inline ටැග් නොසලකා හැර ඇත. ටැග් පදනම් හැඟීම් සඳහා, Orpheus හෝ Bark වැනි ප්රකාශාත්මක ආකෘතිය මාරු.
අභිරුචි උච්චාරණය අර්ථ දක්වන්න (වචනය = උච්චාරණය):
ගැන KhanomTan TTS
KhanomTan TTS was built by Thai NLP contributor Wannaphong Phatthiyaphaibun on top of the YourTTS multilingual VITS architecture, and trained on CC0 and other permissively-licensed Thai corpora including TSync. Where most open TTS models only cover Thai under research-only or non-commercial terms, KhanomTan ships under Apache 2.0, making it a rare commercially-safe choice for the language. It offers a small roster of speaker voices and runs fast at roughly five seconds per generation in around 2GB of VRAM. It fits Thai voiceovers, narration for Thai-language apps, and accessibility tooling where licensing clarity matters.
සඳහා හොඳම: Thai voiceovers, Thai-language content and apps
සියල්ල ගවේශනය කරන්න KhanomTan TTS හඬකෙටියෙන්
- සංවර්ධක
- Wannaphong Phatthiyaphaibun
- බලපත්රය
- Apache 2.0
- සත්ත්වයා
- standard
- වේගය
- fast
- හඬ ක්ලෝන කිරීම
- නෑ
- භාෂා
- Thai
- උපරිම අකුරු
- 500