Tortoise TTS

Tortoise TTS ТТС

A quality-first autoregressive model — slow, but among the most realistic open-source speech available.

Падпісацца Абмежаванне на 5000 знакаў

Захоўваць тэкст у тэгах SSML для дакладнага кантролю:

<speak><prosody rate="slow">Slow speech</prosody></speak>

Тэгі, якія разумее выбраная мадэль - націсніце, каб перанесці іх у тэкст:

Гэтая мадэль чытае звычайны тэкст, таму ўбудаваныя тэгі ігнаруюцца. Для эмоцый, заснаваных на тэгах, пераключыцеся на мадэлі выразнасці, такія як Orpheus або Bark.

Вызначыць уласнае вымаўленне (слова = вымаўленне):

-12 +12
0.5x 2.0x
Свабодны з Piper, VITS, MeloTTS
Створаны вамі гук з' явіцца тут. Выберыце мадэль, увядзіце тэкст і націсніце Стварыць.
Аўдыё паспяхова створанаName
0:00
Сцягнуць гук Сцягнуць.srt Тэрмін дзеяння спасылкі скончыцца праз 24 гадзіны
Любіце TTS.ai? Раскажыце сваім сябрам!

Пра Tortoise TTS

Tortoise TTS, created by James Betker, deliberately trades speed for quality. It is an autoregressive multi-voice system using a DALL-E-inspired architecture, and it produces some of the most realistic synthetic speech in the open-source ecosystem, with excellent prosody and speaker similarity. The name is a nod to its pace: it is noticeably slower than most alternatives, but the payoff is studio-grade output. It supports multiple voices and voice cloning (which benefits from a longer reference, around fifteen seconds), making it a long-standing favorite for audiobooks and premium narration where wait time is acceptable. Tortoise is English-focused and released under the permissive Apache 2.0 license.

Лепшы для: Audiobooks, premium content, quality-first applications

Прагляд усіх Tortoise TTS галасы

Кароткае апісанне

Распрацоўшчык
James Betker
Ліцэнзія
Apache 2.0
Стварыць
premium
Хуткасць
slow
Клонаванне голасу
Так
Мовы
English
Найбольшая колькасць знакаў
2000

Tortoise TTS галасы

Random

English
Прэміум Neutral

Tortoise TTS Частыя пытанні

It is autoregressive and uses a DALL-E-inspired architecture that deliberately prioritizes quality over speed. The trade-off is some of the most realistic open-source speech available, which is why it remains popular for audiobooks despite the wait.

Yes. It supports multi-voice synthesis and voice cloning; results improve with a longer reference, around fifteen seconds of clean audio.

Quality-first applications — audiobooks and premium narration — where its slow but highly realistic output is worth the generation time. It is English-focused and Apache 2.0 licensed.
← Усе галасы