Tortoise TTS

Tortoise TTS 1 - تكنولوجيا المعلومات والاتصالات

A quality-first autoregressive model — slow, but among the most realistic open-source speech available.

انضم 000 5 كلمة

لف نصك في علامات SSML للتحكم الدقيق:

<speak><prosody rate="slow">Slow speech</prosody></speak>

العلامات التي يفهمها النموذج المختار — انقر لإسقاط واحدة في نصك حيث تحصل:

هذا النموذج يقرأ النص العادي، لذلك يتم تجاهل العلامات في السطر. للتعبير عن المشاعر القائمة على العلامات، انتقل إلى نموذج تعبيري مثل أورفيوس أو بارك.

تعريف النطق العادي (كلمة = نطق):

-12 +12
0.5x 2.0x
مجاني مع Piper, VITS, MeloTTS
سيظهر الصوت الذي أنتجته هنا. اختر نموذجاً، وأدخل نصاً، ثم انقر على توليد.
تم توليد الصوت بنجاح
0:00
تنزيل الصوت تنزيل.srt الرابط ينتهي بعد 24 ساعة
المستوى المجاني: الاستخدام الشخصي. ترخيص تجاري من 5 دولارات شهريا
أحب TTS.ai؟ أخبر أصدقائك!

حول Tortoise TTS

Tortoise TTS, created by James Betker, deliberately trades speed for quality. It is an autoregressive multi-voice system using a DALL-E-inspired architecture, and it produces some of the most realistic synthetic speech in the open-source ecosystem, with excellent prosody and speaker similarity. The name is a nod to its pace: it is noticeably slower than most alternatives, but the payoff is studio-grade output. It supports multiple voices and voice cloning (which benefits from a longer reference, around fifteen seconds), making it a long-standing favorite for audiobooks and premium narration where wait time is acceptable. Tortoise is English-focused and released under the permissive Apache 2.0 license.

أفضل لل: Audiobooks, premium content, quality-first applications

تصفح جميع Tortoise TTS الأصوات

لمحة عامة

مطوِّر
James Betker
الترخيص
Apache 2.0
الرتبة
premium
السرعة
slow
استنساخ الصوت
نعم
اللغات
English
الحد الأقصى للحروف
2000

Tortoise TTS الأصوات

Random

English
الأقساط Neutral

Tortoise TTS الأسئلة المتكررة

It is autoregressive and uses a DALL-E-inspired architecture that deliberately prioritizes quality over speed. The trade-off is some of the most realistic open-source speech available, which is why it remains popular for audiobooks despite the wait.

Yes. It supports multi-voice synthesis and voice cloning; results improve with a longer reference, around fifteen seconds of clean audio.

Quality-first applications — audiobooks and premium narration — where its slow but highly realistic output is worth the generation time. It is English-focused and Apache 2.0 licensed.
← جميع الأصوات