Tortoise TTS

Tortoise TTS TTS

A quality-first autoregressive model — slow, but among the most realistic open-source speech available.

Registreeru 5000 tähemärgi piir

SSML-i siltidesse teksti segamine täpseks kontrollimiseks:

<speak><prosody rate="slow">Slow speech</prosody></speak>

Sildid valitud mudelil mõistavad ~ klõpsa ühe kukutamiseks teksti, kus see juhtub:

See mudel loeb lihtsat teksti, nii et sisemisi silte ignoreeritakse. Sildil põhinevate emotsioonide puhul lülituge ekspressiivsele mudelile nagu Orpheus või Bark.

Kohandatud häälduste määramine (sõna = hääldus):

-12 +12
0.5x 2.0x
Tasuta Piper, VITS, MeloTTS
Siin ilmub sinu loodud heli. Vali mudel, sisesta tekst ja klõpsa Genereeri.
Audio genereeritud edukalt
0:00
Audio allalaadimine Lae alla.srt Link aegub 24 tunni pärast.
Tasuta tase: isiklik kasutamine. Äriline litsents alates $5/mo
Armastus TTS.ai?

Info Tortoise TTS

Tortoise TTS, created by James Betker, deliberately trades speed for quality. It is an autoregressive multi-voice system using a DALL-E-inspired architecture, and it produces some of the most realistic synthetic speech in the open-source ecosystem, with excellent prosody and speaker similarity. The name is a nod to its pace: it is noticeably slower than most alternatives, but the payoff is studio-grade output. It supports multiple voices and voice cloning (which benefits from a longer reference, around fifteen seconds), making it a long-standing favorite for audiobooks and premium narration where wait time is acceptable. Tortoise is English-focused and released under the permissive Apache 2.0 license.

Parim: Audiobooks, premium content, quality-first applications

Kõigi sirvimine Tortoise TTS hääled

Põgusalt

Arendaja
James Betker
Litsents
Apache 2.0
Määramistasand
premium
Kiirus
slow
Hääle kloonimine
Jah
Keeled
English
Maks. märgid
2000

Tortoise TTS hääled

Random

English
Premium Neutral

Tortoise TTS TTS (KKK)

It is autoregressive and uses a DALL-E-inspired architecture that deliberately prioritizes quality over speed. The trade-off is some of the most realistic open-source speech available, which is why it remains popular for audiobooks despite the wait.

Yes. It supports multi-voice synthesis and voice cloning; results improve with a longer reference, around fifteen seconds of clean audio.

Quality-first applications — audiobooks and premium narration — where its slow but highly realistic output is worth the generation time. It is English-focused and Apache 2.0 licensed.
← Kõik hääled