Sesame CSM ټي ټي اېس
A 1B conversational speech model that captures natural dialogue timing, turn-taking, and backchannel responses.
د دقیق کنټرول لپاره په SSML نښانونو خپل متن واچوئ:
<speak><prosody rate="slow">Slow speech</prosody></speak>
توري د ټاکل شوي ماډل پوهیږي - کلیک وکړئ چې ستاسو په متن کې یو راښکته کړئ چیرې چې دا پیښیږي:
دا ماډل لوستل ساده متن، نو inline نښانونه په پام کې نه نيول کيږي. د نښان پر بنسټ احساس، د يو څرګند ماډل لکه Orpheus يا Bark بدل.
دوديزه لوستنه پېژندل (ويې = لوستنه):
په اړه Sesame CSM
Sesame CSM (Conversational Speech Model) is a 1-billion-parameter model from Sesame designed specifically for the rhythms of human conversation. Built on a Llama backbone paired with an audio codec, it models turn-taking timing, backchannel responses (the small acknowledgements people make while listening), emotional reactions, and overall conversational flow. The result reads less like read-aloud text and more like a real spoken exchange. It is a natural fit for AI assistants, chatbots, and conversational interfaces where the goal is speech that feels responsive and human. CSM is released under Apache 2.0, and access on TTS.ai requires a Hugging Face token at the model level.
غوره د: AI assistants, chatbots, conversational AI applications
ټول لټول Sesame CSM غږونهپه يوه کتنه کې
- جوړوونکی
- Sesame
- منښتليک
- Apache 2.0
- :د پاڼې نوم
- premium
- چټکتيا
- slow
- غږ کلونول
- نه
- ژبې
- English
- ټولوجګه لوښه
- 500