Tortoise TTS

Tortoise TTS TTS

A quality-first autoregressive model — slow, but among the most realistic open-source speech available.

Підписатися обмеження на 5,000 символів

Переносити ваш текст до міток SSML для точного керування:

<speak><prosody rate="slow">Slow speech</prosody></speak>

Додати позначки емоцій доставки впливу (відносна підтримка model):

Ця модель читає простий текст, отже теґи вбудований буде проігноровано. Для емоцій, заснованих на мітках, перемкніться на модель експресивного вираження, на зразок Orpheus або Bark.

Визначити нетипові вимови (слово = вимова):

-12 +12
0.5x 2.0x
Вільно з Пайпером, VITS, Melotts
Тут з' явиться створений вами звуковий файл. Оберіть модель, введіть текст і натисніть кнопку Створити.
Звук успішно створено
0:00
Звантажити аудіо Звантажити. str Зв' язок закінчується через 24h
Вільний наручник: особисте використання. Комерційна ліцензія з $5/mo
У режимі низького зменшення вільного символу Отримати 200К символів щомісяця або одночасно 100К пачки за 5 доларів
Зробити це вашим власним голосом Клонувати голос за 30 секунд
Любити TTS.ai?

Про програму Tortoise TTS

Tortoise TTS, created by James Betker, deliberately trades speed for quality. It is an autoregressive multi-voice system using a DALL-E-inspired architecture, and it produces some of the most realistic synthetic speech in the open-source ecosystem, with excellent prosody and speaker similarity. The name is a nod to its pace: it is noticeably slower than most alternatives, but the payoff is studio-grade output. It supports multiple voices and voice cloning (which benefits from a longer reference, around fifteen seconds), making it a long-standing favorite for audiobooks and premium narration where wait time is acceptable. Tortoise is English-focused and released under the permissive Apache 2.0 license.

Найкраще для: Audiobooks, premium content, quality-first applications

Перегляд всього Tortoise TTS голоси

На перший погляд

Розробник
James Betker
Ліцензія
Apache 2.0
Тір
premium
Швидкість
slow
Клонування голосів
Так.
Мови
English
Макс. символи
2000

Tortoise TTS голоси

Random

English
Премій Neutral

Tortoise TTS ТЖИ · СО

It is autoregressive and uses a DALL-E-inspired architecture that deliberately prioritizes quality over speed. The trade-off is some of the most realistic open-source speech available, which is why it remains popular for audiobooks despite the wait.

Yes. It supports multi-voice synthesis and voice cloning; results improve with a longer reference, around fifteen seconds of clean audio.

Quality-first applications — audiobooks and premium narration — where its slow but highly realistic output is worth the generation time. It is English-focused and Apache 2.0 licensed.
← Всі голоси