Parler TTS

Parler TTS ТТС

Describe the voice you want in plain English and Parler generates speech matching that description.

Запиши се. за 5000 ограничаване на знака

Опаковка на вашия текст в SSML тагове за точен контрол:

<speak><prosody rate="slow">Slow speech</prosody></speak>

Етикети избрания модел разбира — кликнете за да пуснете един в вашия текст, където се случва:

Този модел чете обикновен текст, така че в линия тагове се пренебрегват. За емоции, базирани на таг, превключете на експресивен модел като Orpheus или Bark.

Определяне на своите изговори (слово = произношение):

-12 +12
0.5x 2.0x
Без пари с Пайпър, ВИТС, Мелотс
Тук ще се появи генерираното ви аудио. Изберете модел, въведете текст и кликнете върху Генериране.
Аудио генерирано успешно
0:00
Изтегляне на аудиото Сваляне.rt. Връзката изтича след 24 часа
Безплатен ступенят: лично използване. Търговска лиценз от $5/мо
Обичай ТТСай, кажи на приятелите си!

За Parler TTS

Parler TTS, developed by Hugging Face, replaces voice presets with natural-language control: instead of picking from a fixed list, you write a description such as "a warm female voice with a slight British accent, speaking slowly and clearly," and the model synthesizes speech to match. This makes it unusually flexible for creative work where you need a specific, custom voice character without recording or cloning anyone. It is an 880M-parameter transformer encoder-decoder trained on roughly 45,000 hours of speech, and it is released under the permissive Apache 2.0 license. Parler is English-focused and best suited to applications that benefit from on-demand, describable voice characteristics.

Най-добро за: Creative applications where you need custom voice characteristics

Преглед на всички Parler TTS гласове

На един поглед.

Разработчик
Hugging Face
Лиценз
Apache 2.0
Ниво на равнището
standard
Скорост
medium
Гласово клониране
Не.
Езици
English
Макс. символи
500

Parler TTS гласове

Default

English
Стандартен Neutral

Parler TTS TTS — Често задавани въпроси

You describe it in natural language — gender, accent, pace, tone, and recording quality — and Parler generates speech matching the description. There are no preset voices to choose from.

Hugging Face. It is an 880M-parameter transformer encoder-decoder trained on around 45,000 hours of speech and released under Apache 2.0.

No. Parler generates a voice from a text description rather than from a reference recording. For cloning a specific voice, use a model like Chatterbox or GPT-SoVITS.
← Всички гласове.