Kokoro ടിടിഎസ്
An 82M-parameter open model from Hexgrad that delivers studio-quality speech at nearly 100x real-time.
കൃത്യമായ നിയന്ത്രണത്തിനായി SSML തൊങ്ങലില് വാചകം പൊതിയുക:
<speak><prosody rate="slow">Slow speech</prosody></speak>
തെരഞ്ഞെടുത്ത മാതൃക മനസ്സിലാക്കുന്നത് ടാഗ് (കുടികള്) :
ഈ മോഡ് സാധാരണ പദാവലി വായിക്കുന്നു, അതുകൊണ്ട് ഇന്ലൈന് തൊങ്ങല് അവഗണിപ്പിക്കുന്നു. ടാഗ് അടിസ്ഥാനപരമായ വികാരങ്ങള്ക്കു് ഓര്ഫിയസ് അല്ലെങ്കില് ബാര്ക് പോലുള്ള ഒരു ചിത്രീകരണ മോഡില് മാറുക.
ഇഷ്ടപ്പെട്ട ഉച്ചാരണം നിര്വ്വചിക്കുക (വാക്ക് = ഉച്ചാരണം):
സംബന്ധിച്ച് Kokoro
Kokoro, built by Hexgrad, is a deliberately tiny 82-million-parameter model that punches far above its size class. It uses a StyleTTS + ISTFTNet architecture and was trained on roughly 1,200 hours of speech, yet generates audio close to 100x faster than real-time on a GPU while staying remarkably natural and expressive. It covers English, Japanese, Chinese, French, Italian, Portuguese, Spanish, and Hindi with a varied set of expressive voicepacks, and supports streaming. The combination of small footprint, low latency, and a permissive license has made Kokoro one of the most popular free models for high-volume and streaming use — it carries the largest share of traffic on TTS.ai.
അതിനു വേണ്ടിയുള്ള ഏറ്റവും നല്ല സ്ഥലം.: High-quality TTS with minimal latency, streaming applications
എല്ലാം പരതുക Kokoro ശബ്ദങ്ങള്ഒരു നോക്കുമ്പോള്
- രചയിതാവു്
- Hexgrad
- അനുമതി
- Apache 2.0
- ടിയെര്
- free
- വേഗത
- fast
- ശബ്ദമിശ്രണോപാധി
- ഇല്ല
- ഭാഷകള്
- English, Japanese, Chinese, French, Italian, Portuguese, Spanish, Hindi
- ഏറ്റവും കൂടിയ ക്യാരക്ടറുകള്
- 500