Parler TTS

Parler TTS TTS

Describe the voice you want in plain English and Parler generates speech matching that description.

Tia sahihi kwa kiwango cha tabia 5,000

Pakua maandishi yako katika tovuti ya SSML kwa ajili ya udhibiti sahihi:

<speak><prosody rate="slow">Slow speech</prosody></speak>

Tag anaelewa mfano unaochaguliwa na unajibu ujumbe huu:

Mfano huu unasomeka maandishi rahisi, kwa hiyo alama za vidole hupuuzwa. Kwa hisia za ndani za watu, geukia kigezo kinachoonesha hisia kama Orfeus au Bark.

Matamshi ya desturi (neno = matamshi):

-12 +12
0.5x 2.0x
Nikiwa huru na Piper, VITS, MelloTTTS
Unaweza kuchagua mfano, maandishi, na kidofo kinachoitwa Genete.
Edio Iliyorekebishwa kwa Mafanikio
0:00
Paketi ya Audio Paketisha.srt Kiungo kinakufa mnamo 24
Safu huru: matumizi ya kibinafsi. Hati ya biashara kutoka dola 5/mo
Fanya hii sauti yako mwenyewe Chokoa sauti kwa sekunde 30
Waeleze rafiki zako kuhusu mapenzi ya TTS.ai?

Habari Parler TTS

Parler TTS, developed by Hugging Face, replaces voice presets with natural-language control: instead of picking from a fixed list, you write a description such as "a warm female voice with a slight British accent, speaking slowly and clearly," and the model synthesizes speech to match. This makes it unusually flexible for creative work where you need a specific, custom voice character without recording or cloning anyone. It is an 880M-parameter transformer encoder-decoder trained on roughly 45,000 hours of speech, and it is released under the permissive Apache 2.0 license. Parler is English-focused and best suited to applications that benefit from on-demand, describable voice characteristics.

Bora kwa: Creative applications where you need custom voice characteristics

Ng'ombe wote Parler TTS sauti

Kutupia jicho

Mbuni
Hugging Face
Lenzi
Apache 2.0
Tier
standard
Mwendo
medium
Kufanyizwa kwa Sauti
Hapana
Lugha
English
Wahusika wa Max
500

Parler TTS sauti

Default

English
Kiwango Neutral

Parler TTS TTS ngumuSTEGAQ

You describe it in natural language — gender, accent, pace, tone, and recording quality — and Parler generates speech matching the description. There are no preset voices to choose from.

Hugging Face. It is an 880M-parameter transformer encoder-decoder trained on around 45,000 hours of speech and released under Apache 2.0.

No. Parler generates a voice from a text description rather than from a reference recording. For cloning a specific voice, use a model like Chatterbox or GPT-SoVITS.
← Sauti zote