Parler TTS

Parler TTS TTS

Describe the voice you want in plain English and Parler generates speech matching that description.

Akaụntụ maka 5,000 akara oghe

Kpọchie ngwe gị n'ime SSML táàbụ̀ maka nlekọta ziri ezi:

<speak><prosody rate="slow">Slow speech</prosody></speak>

Táàbụ̀ nke móòdù ahụ a họọrọ na-aghọta - pịa ka ịkpụga otu n'ime ngwe gị ebe ọ na-eme:

Móòdù a na-agụ ngwe nkịtị, yabụ na a na-ewepụta inline táàbụ̀. Maka táàbụ̀-n'okpuru n'émóòdù, gbanwee ka móòdù na-egosi ihe dịka Orpheus mọọbụ Bark.

Ndesịta okwu emeredịkachọrọ:

-12 +12
0.5x 2.0x
Free na Piper, VITS, MeloTTS
Ọdịdị gị ga-egosipụta ebe a. Họrọ móòdù, tinye ngwe, ma pịa Kewapụta.
Ọdịdị a mepụtala nke ọma
0:00
Bubata ụda Bubata.srt Ndesịta njikọ ahụ ga-agwụ n'ime 24h
Free tier: ojiji onwe onye. Commercial license site na $5/mo
Mee ka ọ bụrụ ụda gị Kloo ụda n'ime sekọnd 30
Ị hụrụ TTS.ai? Kpọtụrụ enyi gị!

_N'ihe banyere Parler TTS

Parler TTS, developed by Hugging Face, replaces voice presets with natural-language control: instead of picking from a fixed list, you write a description such as "a warm female voice with a slight British accent, speaking slowly and clearly," and the model synthesizes speech to match. This makes it unusually flexible for creative work where you need a specific, custom voice character without recording or cloning anyone. It is an 880M-parameter transformer encoder-decoder trained on roughly 45,000 hours of speech, and it is released under the permissive Apache 2.0 license. Parler is English-focused and best suited to applications that benefit from on-demand, describable voice characteristics.

Ọkachasị maka: Creative applications where you need custom voice characteristics

Nlegharịa niile Parler TTS ụda

N'ime nlele

Ńkwádò
Hugging Face
Ikikere
Apache 2.0
Tier
standard
Nhazi
medium
Nhazi ụda
Ọ bụghị
Asụsụ ndị ahụ
English
Ụhara Max
500

Parler TTS ụda

Default

English
Dìfọ́ọ̀ltụ̀ Neutral

Parler TTS TTS - Ajụjụ ndị na-emekarị

You describe it in natural language — gender, accent, pace, tone, and recording quality — and Parler generates speech matching the description. There are no preset voices to choose from.

Hugging Face. It is an 880M-parameter transformer encoder-decoder trained on around 45,000 hours of speech and released under Apache 2.0.

No. Parler generates a voice from a text description rather than from a reference recording. For cloning a specific voice, use a model like Chatterbox or GPT-SoVITS.
← Agụgụala niile