Bark

Bark TTS

Suno's transformer-based text-to-audio model that generates speech plus laughter, sighs, music, and sound effects.

Aanmelden voor 5.000 tekenlimiet

Wrap uw tekst in SSML-tags voor nauwkeurige controle:

<speak><prosody rate="slow">Slow speech</prosody></speak>

Tags het geselecteerde model begrijpt

Dit model leest platte tekst, dus inline tags worden genegeerd. Voor emotie op basis van tags, schakel naar een expressief model zoals Orpheus of Bark.

Definieer aangepaste uitspraaken (woord = uitspraak):

-12 +12
0.5x 2.0x
Gratis met Piper, VITS, MeloTTS
Uw gegenereerde audio zal hier verschijnen. Kies een model, voer tekst in en klik op Genereren.
Audio Generated Succesvol
0:00
Audio downloaden Download.srt Link verloopt in 24 uur
Gratis niveau: persoonlijk gebruik. Commerciële licentie van $5/mo
Hou van TTS.ai? Vertel het je vrienden!

Info Bark

Bark comes from Suno and takes a different approach from most TTS systems: it is a GPT-style transformer trained as a text-to-audio model rather than a pure text-to-speech one. Because it generates raw audio tokens, it can produce nonverbal sounds — laughing, sighing, crying — as well as background music and sound effects alongside the spoken words. It ships with 100+ speaker presets and handles 13+ languages including English, Chinese, French, German, Hindi, Japanese, and Korean. The trade-off is speed and length: at 350M parameters it runs slowly (~15s per clip) and caps at 200 characters, so it shines for short, emotive, creative audio rather than long narration.

Beste voor: Creative audio content, audiobooks with emotion, sound effects

Alles doorbladeren Bark stemmen

In een oogopslag

Ontwikkelaar
Suno
Licentie
MIT
Niveau
standard
Snelheid
slow
Klonen van stemmen
Nee
Talen
English, Chinese, French, German, Hindi, Italian, Japanese, Korean, Polish, Portuguese, Russian, Spanish, Turkish
Max. tekens
200

Bark stemmen

Chinese Speaker 1

Chinese
Standaard Neutral

Chinese Speaker 2

Chinese
Standaard Neutral

English Female 1

English
Standaard Female

English Female 2

English
Standaard Female

English Female 3

English
Standaard Female

English Female 4

English
Standaard Female

English Male 1

English
Standaard Male

English Male 2

English
Standaard Male

English Male 3

English
Standaard Male

English Male 4

English
Standaard Male

English Male 5

English
Standaard Male

English Male 6

English
Standaard Male

French Speaker 1

French
Standaard Neutral

French Speaker 2

French
Standaard Neutral

German Speaker 1

German
Standaard Neutral

German Speaker 2

German
Standaard Neutral

Hindi Speaker 1

Hindi
Standaard Neutral

Italian Speaker 1

Italian
Standaard Neutral

Japanese Speaker 1

Japanese
Standaard Neutral

Japanese Speaker 2

Japanese
Standaard Neutral

Korean Speaker 1

Korean
Standaard Neutral

Korean Speaker 2

Korean
Standaard Neutral

Polish Speaker 1

Polish
Standaard Neutral

Portuguese Speaker 1

Portuguese
Standaard Neutral

Russian Speaker 1

Russian
Standaard Neutral

Spanish Speaker 1

Spanish
Standaard Neutral

Spanish Speaker 2

Spanish
Standaard Neutral

Turkish Speaker 1

Turkish
Standaard Neutral

Bark Veelgestelde vragen

Yes. Bark is a text-to-audio model, so beyond speech it can generate nonverbal cues like laughing, sighing and crying, plus music and background sound effects — one of its defining capabilities.

Yes. Bark is MIT-licensed, which permits commercial use.

Bark caps at 200 characters per request and is on the slower side (around 15 seconds per clip), so it is best suited to short, expressive snippets rather than long-form audio. It does not support voice cloning.
← Alle stemmen