Bark TTS
Suno's transformer-based text-to-audio model that generates speech plus laughter, sighs, music, and sound effects.
Enveloppez votre texte dans des balises SSML pour un contrôle précis :
<speak><prosody rate="slow">Slow speech</prosody></speak>
Mots clés le modèle sélectionné comprend — cliquez pour en déposer un dans votre texte où il se produit:
Ce modèle lit du texte clair, donc les balises en ligne sont ignorées. Pour l'émotion basée sur les tags, passer à un modèle expressif comme Orphée ou Bark.
Définir les prononciations personnalisées (mot = prononciation) :
À propos Bark
Bark comes from Suno and takes a different approach from most TTS systems: it is a GPT-style transformer trained as a text-to-audio model rather than a pure text-to-speech one. Because it generates raw audio tokens, it can produce nonverbal sounds — laughing, sighing, crying — as well as background music and sound effects alongside the spoken words. It ships with 100+ speaker presets and handles 13+ languages including English, Chinese, French, German, Hindi, Japanese, and Korean. The trade-off is speed and length: at 350M parameters it runs slowly (~15s per clip) and caps at 200 characters, so it shines for short, emotive, creative audio rather than long narration.
Meilleur pour: Creative audio content, audiobooks with emotion, sound effects
Tout voir Bark voixEn un coup d'oeil
- Développeur
- Suno
- Licence
- MIT
- Niveau
- standard
- Régime
- slow
- Closonnage de la voix
- Numéro
- Langues
- English, Chinese, French, German, Hindi, Italian, Japanese, Korean, Polish, Portuguese, Russian, Spanish, Turkish
- Personnages maxi
- 200