Tortoise TTS

Tortoise TTS TTS

A quality-first autoregressive model — slow, but among the most realistic open-source speech available.

Izena eman 5.000 karaktereko muga

Itzulbiratu zure testua SSML etiketetan kontrol zehatzagoa lortzeko:

<speak><prosody rate="slow">Slow speech</prosody></speak>

Hautatutako modeloak ulertzen dituen etiketak — egin klik testuan jartzeko:

Eredu honek testu arrunta irakurtzen du, beraz, lerro-barneko etiketei ez zaie jaramonik egiten. Etiketetan oinarritutako emozioetarako, aldatu Orpheus edo Bark bezalako adierazpen-modelo batera.

Definitu ahoskera pertsonalizatuak (hitza = ahoskera):

-12 +12
0.5x 2.0x
Librea Piper, VITS, MeloTTS-ekin
Zure sortutako audioa hemen agertuko da. Aukeratu modelo bat, idatzi testua eta egin klik Sortu botoian.
Audioa behar bezala sortu da
0:00
Deskargatu audioa Deskargatu.srt Esteka 24 ordutan iraungiko da
Librea: erabiltzaile pribatuentzat. Lizentzia komertziala $5/mo-tik
Maite TTS.ai? Esan zure lagunei!

Honi buruz Tortoise TTS

Tortoise TTS, created by James Betker, deliberately trades speed for quality. It is an autoregressive multi-voice system using a DALL-E-inspired architecture, and it produces some of the most realistic synthetic speech in the open-source ecosystem, with excellent prosody and speaker similarity. The name is a nod to its pace: it is noticeably slower than most alternatives, but the payoff is studio-grade output. It supports multiple voices and voice cloning (which benefits from a longer reference, around fifteen seconds), making it a long-standing favorite for audiobooks and premium narration where wait time is acceptable. Tortoise is English-focused and released under the permissive Apache 2.0 license.

Honako hauentzako onena: Audiobooks, premium content, quality-first applications

Arakatu dena Tortoise TTS ahotsak

Begirada batean

Garatzailea
James Betker
Lizentzia
Apache 2.0
Tier
premium
Abiadura
slow
Ahots klonaketa
Bai
Hizkuntzak
English
Gehienezko karaktereak
2000

Tortoise TTS ahotsak

Random

English
Premium Neutral

Tortoise TTS TTS — Galdera ohikoenak

It is autoregressive and uses a DALL-E-inspired architecture that deliberately prioritizes quality over speed. The trade-off is some of the most realistic open-source speech available, which is why it remains popular for audiobooks despite the wait.

Yes. It supports multi-voice synthesis and voice cloning; results improve with a longer reference, around fifteen seconds of clean audio.

Quality-first applications — audiobooks and premium narration — where its slow but highly realistic output is worth the generation time. It is English-focused and Apache 2.0 licensed.
← Ahots guztiak