Parler TTS

Parler TTS TTS

Describe the voice you want in plain English and Parler generates speech matching that description.

Pierakstīties 5000 rakstzīmju limitam

Aplauzt savu tekstu SSML tagus precīzai kontrolei:

<speak><prosody rate="slow">Slow speech</prosody></speak>

Tags izvēlētais modelis saprot — noklikšķiniet, lai iemestu vienu jūsu tekstā, kur tas notiek:

Šis modelis lasa vienkāršu tekstu, tāpēc tiek ignorēti inline tagi. Uz tag-based emocijas, pāriet uz izteiksmīgu modeli, piemēram, Orpheus vai Bark.

Definēt pielāgotu izrunas (vārds = izruna):

-12 +12
0.5x 2.0x
Bez piper, VITS, MeloTTS
Šeit parādīsies jūsu ģenerētais audio. Izvēlieties modeli, ievadiet tekstu un noklikšķiniet ģenerējiet.
Audio veiksmīgi ģenerēts
0:00
Lejupielādēt audio Lejupielādēt.srt Saite beidzas 24h
Bezmaksas līmenis: personīgai lietošanai. Komerclicenci no $5/mo
Mīlestība TTS.ai? Stāsti saviem draugiem!

Par Parler TTS

Parler TTS, developed by Hugging Face, replaces voice presets with natural-language control: instead of picking from a fixed list, you write a description such as "a warm female voice with a slight British accent, speaking slowly and clearly," and the model synthesizes speech to match. This makes it unusually flexible for creative work where you need a specific, custom voice character without recording or cloning anyone. It is an 880M-parameter transformer encoder-decoder trained on roughly 45,000 hours of speech, and it is released under the permissive Apache 2.0 license. Parler is English-focused and best suited to applications that benefit from on-demand, describable voice characteristics.

Labākais: Creative applications where you need custom voice characteristics

Pārlūkot visu Parler TTS balsis

Īsumā

Izstrādātājs
Hugging Face
Licence
Apache 2.0
Līmeņrādis
standard
Ātrums
medium
Balss klonēšana
Valodas
English
Maks. rakstzīmes
500

Parler TTS balsis

Default

English
Standarta Neutral

Parler TTS TTS – FAQ

You describe it in natural language — gender, accent, pace, tone, and recording quality — and Parler generates speech matching the description. There are no preset voices to choose from.

Hugging Face. It is an 880M-parameter transformer encoder-decoder trained on around 45,000 hours of speech and released under Apache 2.0.

No. Parler generates a voice from a text description rather than from a reference recording. For cloning a specific voice, use a model like Chatterbox or GPT-SoVITS.
← Visas balsis