Bark

Bark TTS

Suno's transformer-based text-to-audio model that generates speech plus laughter, sighs, music, and sound effects.

Iscriviti per un limite di 5.000 caratteri

Avvolgi il tuo testo nei tag SSML per un controllo preciso:

<speak><prosody rate="slow">Slow speech</prosody></speak>

Tags il modello selezionato comprende clic su

Questo modello legge testo semplice, quindi i tag inline vengono ignorati. Per le emozioni basate sui tag, passare a un modello espressivo come Orpheus o Bark.

Definire le pronunciazioni personalizzate (parola = pronuncia):

-12 +12
0.5x 2.0x
Gratis con Piper, VITS, MeloTTS
L'audio generato apparirà qui. Scegli un modello, inserisci testo e fai clic su Genera.
Audio generato con successo
0:00
Scarica audio Scarica.srt Link scade in 24 ore
Livello libero: uso personale. Licenza commerciale da $5/mo
Fai di questo la tua voce Clona una voce in 30 secondi
Ti piace TTS.ai? Dillo ai tuoi amici!

Informazioni Bark

Bark comes from Suno and takes a different approach from most TTS systems: it is a GPT-style transformer trained as a text-to-audio model rather than a pure text-to-speech one. Because it generates raw audio tokens, it can produce nonverbal sounds — laughing, sighing, crying — as well as background music and sound effects alongside the spoken words. It ships with 100+ speaker presets and handles 13+ languages including English, Chinese, French, German, Hindi, Japanese, and Korean. The trade-off is speed and length: at 350M parameters it runs slowly (~15s per clip) and caps at 200 characters, so it shines for short, emotive, creative audio rather than long narration.

Meglio per: Creative audio content, audiobooks with emotion, sound effects

Sfoglia tutti Bark voci

A colpo d'occhio

Sviluppatore
Suno
Licenza
MIT
Livello
standard
Velocità
slow
Clonazione vocale
No.
Lingue
English, Chinese, French, German, Hindi, Italian, Japanese, Korean, Polish, Portuguese, Russian, Spanish, Turkish
Caratteri massimi
200

Bark voci

Chinese Speaker 1

Chinese
Standard Neutral

Chinese Speaker 2

Chinese
Standard Neutral

English Female 1

English
Standard Female

English Female 2

English
Standard Female

English Female 3

English
Standard Female

English Female 4

English
Standard Female

English Male 1

English
Standard Male

English Male 2

English
Standard Male

English Male 3

English
Standard Male

English Male 4

English
Standard Male

English Male 5

English
Standard Male

English Male 6

English
Standard Male

French Speaker 1

French
Standard Neutral

French Speaker 2

French
Standard Neutral

German Speaker 1

German
Standard Neutral

German Speaker 2

German
Standard Neutral

Hindi Speaker 1

Hindi
Standard Neutral

Italian Speaker 1

Italian
Standard Neutral

Japanese Speaker 1

Japanese
Standard Neutral

Japanese Speaker 2

Japanese
Standard Neutral

Korean Speaker 1

Korean
Standard Neutral

Korean Speaker 2

Korean
Standard Neutral

Polish Speaker 1

Polish
Standard Neutral

Portuguese Speaker 1

Portuguese
Standard Neutral

Russian Speaker 1

Russian
Standard Neutral

Spanish Speaker 1

Spanish
Standard Neutral

Spanish Speaker 2

Spanish
Standard Neutral

Turkish Speaker 1

Turkish
Standard Neutral

Bark FAQ del TTS

Yes. Bark is a text-to-audio model, so beyond speech it can generate nonverbal cues like laughing, sighing and crying, plus music and background sound effects — one of its defining capabilities.

Yes. Bark is MIT-licensed, which permits commercial use.

Bark caps at 200 characters per request and is on the slower side (around 15 seconds per clip), so it is best suited to short, expressive snippets rather than long-form audio. It does not support voice cloning.
← Tutte le voci