Tortoise TTS

Tortoise TTS TTS

A quality-first autoregressive model — slow, but among the most realistic open-source speech available.

Jijjiirama 5,000 character limit

Daangeessii kitaaba keessan keessaa tag SSML akka itti fayyadamtan:

<speak><prosody rate="slow">Slow speech</prosody></speak>

Tag'oota mo'ellaa filatamee beekuu - cuqaasi akka tokkotti galchiin gara teekstaatti yoo ta'e:

Mo'ellaan kun kitaaba salphaa baraa, kan akka taggaa inniin linjii hin beekkamne. Akka taggaa-based emooshiniitti, mo'ellaa akka Orpheus ykn Bark.

Haalli fuula

-12 +12
0.5x 2.0x
Birrii fi Piper, VITS, MeloTTS
Oduu kee kan uumame yooka'u yooka'u. Suuraa moolaa, galchi kitaaba, fi bu'u Jijjiira.
Audion itti fufuu
0:00
Fuula Oduu Fuula Liqii dhumaa 24 sa'a keessatti
Tarree hin-ga'iin: fayyadama namaatiif. Liiziinii Kominikeeshinii irraa $5/mo
TTS.ai jaallatan? Sochii keessanitti hiika!

Fuulaa Tortoise TTS

Tortoise TTS, created by James Betker, deliberately trades speed for quality. It is an autoregressive multi-voice system using a DALL-E-inspired architecture, and it produces some of the most realistic synthetic speech in the open-source ecosystem, with excellent prosody and speaker similarity. The name is a nod to its pace: it is noticeably slower than most alternatives, but the payoff is studio-grade output. It supports multiple voices and voice cloning (which benefits from a longer reference, around fifteen seconds), making it a long-standing favorite for audiobooks and premium narration where wait time is acceptable. Tortoise is English-focused and released under the permissive Apache 2.0 license.

Fakkeenyaaf: Audiobooks, premium content, quality-first applications

Fuulaa Tortoise TTS Dhaamsa

Akkasumas

Deebi'aa
James Betker
Lizenz
Apache 2.0
Daandiin
premium
Jijjiiramni
slow
Dhaabbilee
Ya
Afaan Oromoo
English
Akkasumas
2000

Tortoise TTS Dhaamsa

Random

English
Premium Neutral

Tortoise TTS TTS — FAQ

It is autoregressive and uses a DALL-E-inspired architecture that deliberately prioritizes quality over speed. The trade-off is some of the most realistic open-source speech available, which is why it remains popular for audiobooks despite the wait.

Yes. It supports multi-voice synthesis and voice cloning; results improve with a longer reference, around fifteen seconds of clean audio.

Quality-first applications — audiobooks and premium narration — where its slow but highly realistic output is worth the generation time. It is English-focused and Apache 2.0 licensed.
← Dhaamsawwan hundaa