Tortoise TTS

Tortoise TTS TTS

A quality-first autoregressive model — slow, but among the most realistic open-source speech available.

Գրանցվել 5000 սանտիմետր սահմանափակում

Ձեր տեքստը SSML տեգերի մեջ տեղադրել ճշգրիտ կառավարման համար.

<speak><prosody rate="slow">Slow speech</prosody></speak>

Ընտրված մոդելի կողմից ընկալվող պիտակներ — սեղմեք դրանք ձեր տեքստում տեղադրելու համար:

Այս մոդելը կարդում է պարզ տեքստը, այնպես որ ներառված տեքստերը անտեսվում են։ Տեքստերի վրա հիմնված էմոցիաների համար փոխեք արտահայտիչ մոդել, ինչպես Orpheus կամ Bark։

Որոշել սեփական արտասանությունը (բառ = արտասանություն):

-12 +12
0.5x 2.0x
Ազատ Piper, VITS, MeloTTS-ով
Այստեղ կհայտնվի ձեր ստեղծած ձայնը։ Ընտրեք մոդել, ներդրեք տեքստ և սեղմեք Ծնվել։
Ավտոմատ ձայնագրում
0:00
Տեղադրել ձայնային Տեղադրել.srt Հղումն ավարտվում է 24 ժամ անց
Ֆրանսիա: Ֆրանսիայի ազգային ռադիո: Ֆրանսիայի ազգային ռադիո. Կազմակերպական լիցենզիա $5/ամս
Սիրում եք TTS.ai-ն? Պատմեք ձեր ընկերներին։

Ընդհանուր Tortoise TTS

Tortoise TTS, created by James Betker, deliberately trades speed for quality. It is an autoregressive multi-voice system using a DALL-E-inspired architecture, and it produces some of the most realistic synthetic speech in the open-source ecosystem, with excellent prosody and speaker similarity. The name is a nod to its pace: it is noticeably slower than most alternatives, but the payoff is studio-grade output. It supports multiple voices and voice cloning (which benefits from a longer reference, around fifteen seconds), making it a long-standing favorite for audiobooks and premium narration where wait time is acceptable. Tortoise is English-focused and released under the permissive Apache 2.0 license.

Լավագույնը: Audiobooks, premium content, quality-first applications

Ընթերցել բոլորը Tortoise TTS ձայներ

Հակառակորդի դիրքը

Հեղինակ
James Betker
Լիցենզիա
Apache 2.0
Դադար
premium
արագություն
slow
Ձայնի կլոնավորում
Այո
Լեզուներ
English
Օգտագործված ռեժիմ
2000

Tortoise TTS ձայներ

Random

English
Պրեմիում Neutral

Tortoise TTS TTS - Հաճախ տրվող հարցեր

It is autoregressive and uses a DALL-E-inspired architecture that deliberately prioritizes quality over speed. The trade-off is some of the most realistic open-source speech available, which is why it remains popular for audiobooks despite the wait.

Yes. It supports multi-voice synthesis and voice cloning; results improve with a longer reference, around fifteen seconds of clean audio.

Quality-first applications — audiobooks and premium narration — where its slow but highly realistic output is worth the generation time. It is English-focused and Apache 2.0 licensed.
← Բոլոր ձայները