Parler TTS

Parler TTS TTS

Describe the voice you want in plain English and Parler generates speech matching that description.

Inscrivez-vous pour la limite de 5 000 caractères

Enveloppez votre texte dans des balises SSML pour un contrôle précis :

<speak><prosody rate="slow">Slow speech</prosody></speak>

Mots clés le modèle sélectionné comprend — cliquez pour en déposer un dans votre texte où il se produit:

Ce modèle lit du texte clair, donc les balises en ligne sont ignorées. Pour l'émotion basée sur les tags, passer à un modèle expressif comme Orphée ou Bark.

Définir les prononciations personnalisées (mot = prononciation) :

-12 +12
0.5x 2.0x
Gratuit avec Piper, VITS, MeloTTS
Votre audio généré apparaîtra ici. Choisissez un modèle, entrez le texte et cliquez sur Générer.
Production audio réussie
0:00
Télécharger l'audio Télécharger.srt Lien expire en 24h
Niveau gratuit: usage personnel. Licence commerciale à partir de 5 $/mois
Vous aimez TTS.ai ? Parlez-en à vos amis !

À propos Parler TTS

Parler TTS, developed by Hugging Face, replaces voice presets with natural-language control: instead of picking from a fixed list, you write a description such as "a warm female voice with a slight British accent, speaking slowly and clearly," and the model synthesizes speech to match. This makes it unusually flexible for creative work where you need a specific, custom voice character without recording or cloning anyone. It is an 880M-parameter transformer encoder-decoder trained on roughly 45,000 hours of speech, and it is released under the permissive Apache 2.0 license. Parler is English-focused and best suited to applications that benefit from on-demand, describable voice characteristics.

Meilleur pour: Creative applications where you need custom voice characteristics

Tout voir Parler TTS voix

En un coup d'oeil

Développeur
Hugging Face
Licence
Apache 2.0
Niveau
standard
Régime
medium
Closonnage de la voix
Numéro
Langues
English
Personnages maxi
500

Parler TTS voix

Default

English
Norme Neutral

Parler TTS TTS — FAQ

You describe it in natural language — gender, accent, pace, tone, and recording quality — and Parler generates speech matching the description. There are no preset voices to choose from.

Hugging Face. It is an 880M-parameter transformer encoder-decoder trained on around 45,000 hours of speech and released under Apache 2.0.

No. Parler generates a voice from a text description rather than from a reference recording. For cloning a specific voice, use a model like Chatterbox or GPT-SoVITS.
← Toutes les voix