Parler TTS

Parler TTS نووسراو بۆ بیستراو

Describe the voice you want in plain English and Parler generates speech matching that description.

تۆماربکە بۆ ٥٠٠٠ هێما

نوسراوەکەت بگۆڕە بۆ تگەکانی SSML بۆ کۆنتڕۆڵی ڕاستەقینە:

<speak><prosody rate="slow">Slow speech</prosody></speak>

تگەکان کە مۆدێلی هەڵبژێردراو تێیدەگات - بکەرەوە بۆ ئەوەی یەکێکیان بخەیتە ناو دەقەکەتەوە کە تیایدا ڕوودەدات:

ئەم مۆدێلە نوسراوێکی ئاسایی دەخوێنێتەوە، بۆیە تاگەکانی ناو ڕستە پشتگوێ دەخرێت. بۆ هەستێکی لەسەر بنەمای تاگ، بگەڕێ بۆ مۆدێلێکی دەربڕین وەک ئۆرفیۆس یان بارک.

پێناسەکردنی دەنگی خۆت (وشە = دەنگی):

-12 +12
0.5x 2.0x
بەبێ پارە لەگەڵ Piper, VITS, MeloTTS
دەنگی دروستکراوت لێرەدا دەردەکەوێت. مۆدێلێک هەڵبژێرە، نوسراوێک دابنێ، پاشان کلیک بکە لەسەر دروستکردن.
دەنگ بە سەرکەوتن دروست کرا
0:00
دابەزاندنی دەنگ دابەزاندن پەیوەستەکە دوای ٢٤ کاتژمێر کۆتایی دێت
پلەی ئازاد: بەکارهێنانی تایبەتی. مۆڵەتی بازرگانی لە ٥$/ مانگ
خۆشت دەوێت TTS.ai؟ بە هاوڕێکانت بڵێ!

دەربارەی Parler TTS

Parler TTS, developed by Hugging Face, replaces voice presets with natural-language control: instead of picking from a fixed list, you write a description such as "a warm female voice with a slight British accent, speaking slowly and clearly," and the model synthesizes speech to match. This makes it unusually flexible for creative work where you need a specific, custom voice character without recording or cloning anyone. It is an 880M-parameter transformer encoder-decoder trained on roughly 45,000 hours of speech, and it is released under the permissive Apache 2.0 license. Parler is English-focused and best suited to applications that benefit from on-demand, describable voice characteristics.

باشترین بۆ: Creative applications where you need custom voice characteristics

سەردانی هەموویان بکە Parler TTS دەنگی

چاوپێکەوتن

پەرەپێدەر
Hugging Face
مۆڵەتی بەکارھێنەر
Apache 2.0
یه‌مه‌ن
standard
خێرایی
medium
دووبارە دروستکردنی دەنگی
نەخێر
زمان
English
زۆرترین پیت
500

Parler TTS دەنگی

Default

English
ستاندارد Neutral

Parler TTS پرسیاری زۆر کراوە

You describe it in natural language — gender, accent, pace, tone, and recording quality — and Parler generates speech matching the description. There are no preset voices to choose from.

Hugging Face. It is an 880M-parameter transformer encoder-decoder trained on around 45,000 hours of speech and released under Apache 2.0.

No. Parler generates a voice from a text description rather than from a reference recording. For cloning a specific voice, use a model like Chatterbox or GPT-SoVITS.
← هەموو دەنگەکان