Tortoise TTS

Tortoise TTS ТТС

A quality-first autoregressive model — slow, but among the most realistic open-source speech available.

Каттоо 5000 символго чейин

Текстти SSML тегдерине өткөрүп берүү:

<speak><prosody rate="slow">Slow speech</prosody></speak>

Белгилер, тандалган модель түшүнөт — текстке бирин коюу үчүн, аны чыкылдатыңыз:

Бул модель жөнөкөй текстти окуйт, ошондуктан тексттеги тегдер эске алынбайт. Тегдер менен эмоцияларды жаратууда Orpheus же Bark сыяктуу экспрессивдүү моделдерге өтүү керек.

Өзгөчө сүйлөмдөрдү аныктоо (сөз = сүйлөм):

-12 +12
0.5x 2.0x
Piper, VITS, MeloTTS менен акысыз
Сиздин түзүлгөн аудио файлыңыз бул жерде пайда болот. Модель тандап, текстти киргизип, Жаңылоо баскычын басыңыз.
Аудио ийгиликтүү түзүлгөн
0:00
Аудиону жүктөп алуу .srt жүктөп алуу Ссылканын мөөнөтү 24 сааттан кийин аяктайт
Free level: жеке колдонуу. Коммерциялык лицензия $5/айдан
Бул үндү өзүңүздүн үнүңүзгө айландырыңыз 30 секундда үндү клондоо
TTS.ai сизге жактыбы? Досторуңузга айтып коюңуз!

Маалымат Tortoise TTS

Tortoise TTS, created by James Betker, deliberately trades speed for quality. It is an autoregressive multi-voice system using a DALL-E-inspired architecture, and it produces some of the most realistic synthetic speech in the open-source ecosystem, with excellent prosody and speaker similarity. The name is a nod to its pace: it is noticeably slower than most alternatives, but the payoff is studio-grade output. It supports multiple voices and voice cloning (which benefits from a longer reference, around fifteen seconds), making it a long-standing favorite for audiobooks and premium narration where wait time is acceptable. Tortoise is English-focused and released under the permissive Apache 2.0 license.

Эң жакшысы: Audiobooks, premium content, quality-first applications

Баарын кароо Tortoise TTS үн

Бир көз салуу

Жазуучу
James Betker
Лицензия
Apache 2.0
Тигр
premium
Жылдамдыгы
slow
Сөздү клондоо
Ооба
Тилдер
English
Макс. символдор
2000

Tortoise TTS үн

Random

English
Премиум Neutral

Tortoise TTS ТТС — Жогорудагы суроолор

It is autoregressive and uses a DALL-E-inspired architecture that deliberately prioritizes quality over speed. The trade-off is some of the most realistic open-source speech available, which is why it remains popular for audiobooks despite the wait.

Yes. It supports multi-voice synthesis and voice cloning; results improve with a longer reference, around fifteen seconds of clean audio.

Quality-first applications — audiobooks and premium narration — where its slow but highly realistic output is worth the generation time. It is English-focused and Apache 2.0 licensed.
← Бардык үн