Parler TTS

Parler TTS TTS

Describe the voice you want in plain English and Parler generates speech matching that description.

Înscrie-te pentru limitele de 5000 de caractere

Întoarceți textul în etichetele SSML pentru un control precis:

<speak><prosody rate="slow">Slow speech</prosody></speak>

Etichetele modelului selectat înțeleg — click pentru a lăsa unul în textul tău unde se întâmplă:

Acest model citește textul simplu, astfel încât etichetele inline sunt ignorate. Pentru emoții bazate pe tag, schimbați la un model expresiv cum ar fi Orpheus sau Bark.

Definiți pronunțiare personalizată (cuvânt = pronunție):

-12 +12
0.5x 2.0x
Gratuit cu Piper, VITS, MeloTTS
Audio generat va apărea aici. Alegeți un model, introduceți text și faceți clic pe Generați.
Audio generat cu succes
0:00
Descarcă audio Descărcare.srt Legătura expiră în 24 ore
Gratuit: utilizare personală. Licență comercială de la 5$/mo
Spune-i prietenilor tăi!

Despre Parler TTS

Parler TTS, developed by Hugging Face, replaces voice presets with natural-language control: instead of picking from a fixed list, you write a description such as "a warm female voice with a slight British accent, speaking slowly and clearly," and the model synthesizes speech to match. This makes it unusually flexible for creative work where you need a specific, custom voice character without recording or cloning anyone. It is an 880M-parameter transformer encoder-decoder trained on roughly 45,000 hours of speech, and it is released under the permissive Apache 2.0 license. Parler is English-focused and best suited to applications that benefit from on-demand, describable voice characteristics.

Cel mai bun pentru: Creative applications where you need custom voice characteristics

Navigați toate Parler TTS voci

La o privire

Dezvoltator
Hugging Face
Licență
Apache 2.0
Nivel
standard
Viteză
medium
Clonarea vocală
Nu.
Limbi
English
Caractere maxime
500

Parler TTS voci

Default

English
Standard Neutral

Parler TTS TTS – FAQ

You describe it in natural language — gender, accent, pace, tone, and recording quality — and Parler generates speech matching the description. There are no preset voices to choose from.

Hugging Face. It is an 880M-parameter transformer encoder-decoder trained on around 45,000 hours of speech and released under Apache 2.0.

No. Parler generates a voice from a text description rather than from a reference recording. For cloning a specific voice, use a model like Chatterbox or GPT-SoVITS.
← Toate vocile