Tortoise TTS TTS
A quality-first autoregressive model — slow, but among the most realistic open-source speech available.
Wrap uw tekst in SSML-tags voor nauwkeurige controle:
<speak><prosody rate="slow">Slow speech</prosody></speak>
Tags het geselecteerde model begrijpt
Dit model leest platte tekst, dus inline tags worden genegeerd. Voor emotie op basis van tags, schakel naar een expressief model zoals Orpheus of Bark.
Definieer aangepaste uitspraaken (woord = uitspraak):
Info Tortoise TTS
Tortoise TTS, created by James Betker, deliberately trades speed for quality. It is an autoregressive multi-voice system using a DALL-E-inspired architecture, and it produces some of the most realistic synthetic speech in the open-source ecosystem, with excellent prosody and speaker similarity. The name is a nod to its pace: it is noticeably slower than most alternatives, but the payoff is studio-grade output. It supports multiple voices and voice cloning (which benefits from a longer reference, around fifteen seconds), making it a long-standing favorite for audiobooks and premium narration where wait time is acceptable. Tortoise is English-focused and released under the permissive Apache 2.0 license.
Beste voor: Audiobooks, premium content, quality-first applications
Alles doorbladeren Tortoise TTS stemmenIn een oogopslag
- Ontwikkelaar
- James Betker
- Licentie
- Apache 2.0
- Niveau
- premium
- Snelheid
- slow
- Klonen van stemmen
- Ja.
- Talen
- English
- Max. tekens
- 2000