Parler TTS

Parler TTS TTS

Describe the voice you want in plain English and Parler generates speech matching that description.

رجسٽر ٿيو 5000 ڪارڪنن جي حد

صحيح ڪنٽرول لاءِ پنھنجو متن SSML ٽيگ ۾ ويڙھيو:

<speak><prosody rate="slow">Slow speech</prosody></speak>

ٽيگ جيڪي چونڊيل ماڊل سمجھي ٿو - هڪ کي پنھنجي متن ۾ جتي ٿئي ٿو ڦيريڻ لاءِ ڪلڪ ڪريو:

هي ماڊل عام متن پڙهندو آھي، تنھنڪري لاٽ ۾ ٽيگ نظرانداز ڪيا ويندا آھن. ٽيگ تي ٻڌل احساسن لاءِ، ھڪ اظهاري ماڊل وانگر Orpheus يا Bark تي تبديل ڪريو.

پنھنجو آواز بيان ڪريو (شيء = آواز):

-12 +12
0.5x 2.0x
پيپر، VITS، MeloTTS سان مفت
پنھنجو ٺاھيل آڊيو اتي نظر ايندو. ھڪ ماڊل چونڊيو، متن داخل ڪريو ۽ ٺاھڻ دٻايو.
آڊيو ڪاميابي سان ٺاهيو ويو
0:00
آڊيو ڊائون لوڊ ڪريو ڊائون لوڊ لنڪ 24 ڪلاڪن ۾ ختم ٿيندو
مفت ٽائير: ذاتي استعمال. $5/مھينن کان تجارتي لائسنس
TTS.ai کي پيارو آهي؟ پنھنجن دوستن کي چئو!

بابت Parler TTS

Parler TTS, developed by Hugging Face, replaces voice presets with natural-language control: instead of picking from a fixed list, you write a description such as "a warm female voice with a slight British accent, speaking slowly and clearly," and the model synthesizes speech to match. This makes it unusually flexible for creative work where you need a specific, custom voice character without recording or cloning anyone. It is an 880M-parameter transformer encoder-decoder trained on roughly 45,000 hours of speech, and it is released under the permissive Apache 2.0 license. Parler is English-focused and best suited to applications that benefit from on-demand, describable voice characteristics.

بهترين: Creative applications where you need custom voice characteristics

سڀ لکو Parler TTS آواز

هڪ نظر ۾

ڊيولپر
Hugging Face
لائسنس
Apache 2.0
جانور
standard
رفتار
medium
آواز جو کلون
نه
ٻوليون
English
وڌيڪ نشان
500

Parler TTS آواز

Default

English
معياري Neutral

Parler TTS TTS - پڇا ڳاڇا

You describe it in natural language — gender, accent, pace, tone, and recording quality — and Parler generates speech matching the description. There are no preset voices to choose from.

Hugging Face. It is an 880M-parameter transformer encoder-decoder trained on around 45,000 hours of speech and released under Apache 2.0.

No. Parler generates a voice from a text description rather than from a reference recording. For cloning a specific voice, use a model like Chatterbox or GPT-SoVITS.
← سڀ آواز