Tortoise TTS

Tortoise TTS TTS

A quality-first autoregressive model — slow, but among the most realistic open-source speech available.

Inscríbete para el límite de 5.000 caracteres

Envuelva su texto en etiquetas SSML para un control preciso:

<speak><prosody rate="slow">Slow speech</prosody></speak>

Etiquetas el modelo seleccionado entiende — haga clic para soltar uno en su texto donde sucede:

Este modelo lee texto plano, por lo que las etiquetas en línea son ignoradas. Para la emoción basada en etiquetas, cambie a un modelo expresivo como Orfeo o Bark.

Definir pronunciaciones personalizadas (palabra = pronunciación):

-12 +12
0.5x 2.0x
Libre con Piper, VITS, MeloTTS
Su audio generado aparecerá aquí. Elija un modelo, introduzca texto y haga clic en Generar.
Audio generado con éxito
0:00
Descargar audio Descargar.srt Enlace expira en 24h
Nivel libre: uso personal. Licencia comercial desde $5/mes
Haz de esto tu propia voz Clonar una voz en 30 segundos
¿Te gusta TTS.ai? ¡Cuéntaselo a tus amigos!

Acerca de Tortoise TTS

Tortoise TTS, created by James Betker, deliberately trades speed for quality. It is an autoregressive multi-voice system using a DALL-E-inspired architecture, and it produces some of the most realistic synthetic speech in the open-source ecosystem, with excellent prosody and speaker similarity. The name is a nod to its pace: it is noticeably slower than most alternatives, but the payoff is studio-grade output. It supports multiple voices and voice cloning (which benefits from a longer reference, around fifteen seconds), making it a long-standing favorite for audiobooks and premium narration where wait time is acceptable. Tortoise is English-focused and released under the permissive Apache 2.0 license.

Lo mejor para: Audiobooks, premium content, quality-first applications

Examinar todo Tortoise TTS voces

De un vistazo

Desarrollador
James Betker
Licencia
Apache 2.0
Nivel
premium
Velocidad
slow
Clonación de voz
Idiomas
English
Máx. caracteres
2000

Tortoise TTS voces

Random

English
Prima Neutral

Tortoise TTS TTS — Preguntas más frecuentes

It is autoregressive and uses a DALL-E-inspired architecture that deliberately prioritizes quality over speed. The trade-off is some of the most realistic open-source speech available, which is why it remains popular for audiobooks despite the wait.

Yes. It supports multi-voice synthesis and voice cloning; results improve with a longer reference, around fifteen seconds of clean audio.

Quality-first applications — audiobooks and premium narration — where its slow but highly realistic output is worth the generation time. It is English-focused and Apache 2.0 licensed.
← Todas las voces