Parler TTS

Parler TTS TTS

Describe the voice you want in plain English and Parler generates speech matching that description.

Inscríbete límite de 5. 000 caracteres

Incluír o texto en etiquetas SSML para un control preciso:

<speak><prosody rate="slow">Slow speech</prosody></speak>

Etiquetas que o modelo escollido entende - prema para deixar unha no texto onde ocorre:

Este modelo le texto simple, polo que se ignoran as etiquetas inline. Para emocións baseadas en etiquetas, cambie a un modelo expresivo como Orpheus ou Bark.

Definir pronunciacións personalizadas (palabra = pronunciación):

-12 +12
0.5x 2.0x
Libre con Piper, VITS, MeloTTS
O son xerado aparecerá aquí. Escolla un modelo, introduza o texto e prema Xerar.
O son xerou correctamente
0:00
Obter o son Obter.srt A ligazón caduca en 24 horas
Nivel libre: uso persoal. Licenza comercial desde $5/mes
Encántalle TTS.ai? Cóntallo aos teus amigos!

Acerca de Parler TTS

Parler TTS, developed by Hugging Face, replaces voice presets with natural-language control: instead of picking from a fixed list, you write a description such as "a warm female voice with a slight British accent, speaking slowly and clearly," and the model synthesizes speech to match. This makes it unusually flexible for creative work where you need a specific, custom voice character without recording or cloning anyone. It is an 880M-parameter transformer encoder-decoder trained on roughly 45,000 hours of speech, and it is released under the permissive Apache 2.0 license. Parler is English-focused and best suited to applications that benefit from on-demand, describable voice characteristics.

Mellor para: Creative applications where you need custom voice characteristics

Examinar todo Parler TTS voces

De un vistazo

Desenvolvente
Hugging Face
Licenza
Apache 2.0
Tier
standard
Velocidade
medium
Clonaxe de voz
Non
Linguas
English
Caracteres máximos
500

Parler TTS voces

Default

English
Estándar Neutral

Parler TTS TTS - FAQ

You describe it in natural language — gender, accent, pace, tone, and recording quality — and Parler generates speech matching the description. There are no preset voices to choose from.

Hugging Face. It is an 880M-parameter transformer encoder-decoder trained on around 45,000 hours of speech and released under Apache 2.0.

No. Parler generates a voice from a text description rather than from a reference recording. For cloning a specific voice, use a model like Chatterbox or GPT-SoVITS.
← Todas as voces