Bark

Bark TTS

Suno's transformer-based text-to-audio model that generates speech plus laughter, sighs, music, and sound effects.

Inscrivez-vous pour la limite de 5 000 caractères

Enveloppez votre texte dans des balises SSML pour un contrôle précis :

<speak><prosody rate="slow">Slow speech</prosody></speak>

Mots clés le modèle sélectionné comprend — cliquez pour en déposer un dans votre texte où il se produit:

Ce modèle lit du texte clair, donc les balises en ligne sont ignorées. Pour l'émotion basée sur les tags, passer à un modèle expressif comme Orphée ou Bark.

Définir les prononciations personnalisées (mot = prononciation) :

-12 +12
0.5x 2.0x
Gratuit avec Piper, VITS, MeloTTS
Votre audio généré apparaîtra ici. Choisissez un modèle, entrez le texte et cliquez sur Générer.
Production audio réussie
0:00
Télécharger l'audio Télécharger.srt Lien expire en 24h
Niveau gratuit: usage personnel. Licence commerciale à partir de 5 $/mois
Vous aimez TTS.ai ? Parlez-en à vos amis !

À propos Bark

Bark comes from Suno and takes a different approach from most TTS systems: it is a GPT-style transformer trained as a text-to-audio model rather than a pure text-to-speech one. Because it generates raw audio tokens, it can produce nonverbal sounds — laughing, sighing, crying — as well as background music and sound effects alongside the spoken words. It ships with 100+ speaker presets and handles 13+ languages including English, Chinese, French, German, Hindi, Japanese, and Korean. The trade-off is speed and length: at 350M parameters it runs slowly (~15s per clip) and caps at 200 characters, so it shines for short, emotive, creative audio rather than long narration.

Meilleur pour: Creative audio content, audiobooks with emotion, sound effects

Tout voir Bark voix

En un coup d'oeil

Développeur
Suno
Licence
MIT
Niveau
standard
Régime
slow
Closonnage de la voix
Numéro
Langues
English, Chinese, French, German, Hindi, Italian, Japanese, Korean, Polish, Portuguese, Russian, Spanish, Turkish
Personnages maxi
200

Bark voix

Chinese Speaker 1

Chinese
Norme Neutral

Chinese Speaker 2

Chinese
Norme Neutral

English Female 1

English
Norme Female

English Female 2

English
Norme Female

English Female 3

English
Norme Female

English Female 4

English
Norme Female

English Male 1

English
Norme Male

English Male 2

English
Norme Male

English Male 3

English
Norme Male

English Male 4

English
Norme Male

English Male 5

English
Norme Male

English Male 6

English
Norme Male

French Speaker 1

French
Norme Neutral

French Speaker 2

French
Norme Neutral

German Speaker 1

German
Norme Neutral

German Speaker 2

German
Norme Neutral

Hindi Speaker 1

Hindi
Norme Neutral

Italian Speaker 1

Italian
Norme Neutral

Japanese Speaker 1

Japanese
Norme Neutral

Japanese Speaker 2

Japanese
Norme Neutral

Korean Speaker 1

Korean
Norme Neutral

Korean Speaker 2

Korean
Norme Neutral

Polish Speaker 1

Polish
Norme Neutral

Portuguese Speaker 1

Portuguese
Norme Neutral

Russian Speaker 1

Russian
Norme Neutral

Spanish Speaker 1

Spanish
Norme Neutral

Spanish Speaker 2

Spanish
Norme Neutral

Turkish Speaker 1

Turkish
Norme Neutral

Bark TTS — FAQ

Yes. Bark is a text-to-audio model, so beyond speech it can generate nonverbal cues like laughing, sighing and crying, plus music and background sound effects — one of its defining capabilities.

Yes. Bark is MIT-licensed, which permits commercial use.

Bark caps at 200 characters per request and is on the slower side (around 15 seconds per clip), so it is best suited to short, expressive snippets rather than long-form audio. It does not support voice cloning.
← Toutes les voix