Bark

Bark TTS

Suno's transformer-based text-to-audio model that generates speech plus laughter, sighs, music, and sound effects.

Inscrever-se para o limite de 5000 caracteres

Envolva o seu texto em tags SSML para controle preciso:

<speak><prosody rate="slow">Slow speech</prosody></speak>

Etiquetas o modelo selecionado entende — clique para soltar um para o seu texto onde acontece:

Este modelo lê texto simples, por isso as etiquetas inline são ignoradas. Para emoção baseada em tags, mude para um modelo expressivo como Orpheus ou Bark.

Definir pronúncias personalizadas (palavra = pronúncia):

-12 +12
0.5x 2.0x
Grátis com Piper, VITS, MeloTTS
Seu áudio gerado aparecerá aqui. Escolha um modelo, introduza texto e clique em Gerar.
O áudio gerado com sucesso
0:00
Baixe áudio Baixar.srt A ligação expira em 24h
Gratuito nível: uso pessoal. Licença comercial a partir de $5/mo
Gosta do TTS.ai? Conte aos seus amigos!

Sobre Bark

Bark comes from Suno and takes a different approach from most TTS systems: it is a GPT-style transformer trained as a text-to-audio model rather than a pure text-to-speech one. Because it generates raw audio tokens, it can produce nonverbal sounds — laughing, sighing, crying — as well as background music and sound effects alongside the spoken words. It ships with 100+ speaker presets and handles 13+ languages including English, Chinese, French, German, Hindi, Japanese, and Korean. The trade-off is speed and length: at 350M parameters it runs slowly (~15s per clip) and caps at 200 characters, so it shines for short, emotive, creative audio rather than long narration.

Melhor para: Creative audio content, audiobooks with emotion, sound effects

Procurar todos Bark vozes

De uma olhada

Desenvolvedor
Suno
Licença
MIT
Tier
standard
Velocidade
slow
Clonagem de voz
Não
Línguas
English, Chinese, French, German, Hindi, Italian, Japanese, Korean, Polish, Portuguese, Russian, Spanish, Turkish
Número máximo de caracteres
200

Bark vozes

Chinese Speaker 1

Chinese
Norma Neutral

Chinese Speaker 2

Chinese
Norma Neutral

English Female 1

English
Norma Female

English Female 2

English
Norma Female

English Female 3

English
Norma Female

English Female 4

English
Norma Female

English Male 1

English
Norma Male

English Male 2

English
Norma Male

English Male 3

English
Norma Male

English Male 4

English
Norma Male

English Male 5

English
Norma Male

English Male 6

English
Norma Male

French Speaker 1

French
Norma Neutral

French Speaker 2

French
Norma Neutral

German Speaker 1

German
Norma Neutral

German Speaker 2

German
Norma Neutral

Hindi Speaker 1

Hindi
Norma Neutral

Italian Speaker 1

Italian
Norma Neutral

Japanese Speaker 1

Japanese
Norma Neutral

Japanese Speaker 2

Japanese
Norma Neutral

Korean Speaker 1

Korean
Norma Neutral

Korean Speaker 2

Korean
Norma Neutral

Polish Speaker 1

Polish
Norma Neutral

Portuguese Speaker 1

Portuguese
Norma Neutral

Russian Speaker 1

Russian
Norma Neutral

Spanish Speaker 1

Spanish
Norma Neutral

Spanish Speaker 2

Spanish
Norma Neutral

Turkish Speaker 1

Turkish
Norma Neutral

Bark TTS — FAQ

Yes. Bark is a text-to-audio model, so beyond speech it can generate nonverbal cues like laughing, sighing and crying, plus music and background sound effects — one of its defining capabilities.

Yes. Bark is MIT-licensed, which permits commercial use.

Bark caps at 200 characters per request and is on the slower side (around 15 seconds per clip), so it is best suited to short, expressive snippets rather than long-form audio. It does not support voice cloning.
← Todas as vozes