Tortoise TTS TTS
A quality-first autoregressive model — slow, but among the most realistic open-source speech available.
Wrap ou tèks nan SSML tags pou presizyon kontwòl:
<speak><prosody rate="slow">Slow speech</prosody></speak>
Tags ke modèl la chwazi konprann — klike pou mete yon nan tèks ou kote li rive:
Modèl sa a li tèks senp, se poutèt sa atik ki nan liy yo pa pran an kont. Pou efè ki baze sou atik, chanje pou yon modèl ekspresyon tankou Orpheus oswa Bark.
Define prononciations Custom (mot = prononciation):
Atik Tortoise TTS
Tortoise TTS, created by James Betker, deliberately trades speed for quality. It is an autoregressive multi-voice system using a DALL-E-inspired architecture, and it produces some of the most realistic synthetic speech in the open-source ecosystem, with excellent prosody and speaker similarity. The name is a nod to its pace: it is noticeably slower than most alternatives, but the payoff is studio-grade output. It supports multiple voices and voice cloning (which benefits from a longer reference, around fifteen seconds), making it a long-standing favorite for audiobooks and premium narration where wait time is acceptable. Tortoise is English-focused and released under the permissive Apache 2.0 license.
Pi bon pou: Audiobooks, premium content, quality-first applications
Navigue tout Tortoise TTS VoyYon ti gade
- Pwogramè
- James Betker
- Lisans
- Apache 2.0
- Nivo
- premium
- Vitès
- slow
- Klonaj vwa
- Wi
- Lang
- English
- Karakteris maksimòm
- 2000