Bark TTS
Suno's transformer-based text-to-audio model that generates speech plus laughter, sighs, music, and sound effects.
Whāriki i tōna kupu i roto i ngā tohu SSML mō te whakahaere tika:
<speak><prosody rate="slow">Slow speech</prosody></speak>
E mōhio ana ngā tohu ki te tauira i kōwhiria - ka kōwhiria kia whakawātea tētahi ki roto i tōna kupu i reira ka puta ai:
Ka pānui tēnei tauira i te kupu noa, nā reira ka whakakāhoretia ngā tohu ā-waitara. Mō te āhua o te tohu-taihi, ka huri ki tētahi tauira whakamārama pēnei i a Orpheus, Bark rānei.
Ka tautuhia ngā tohutohu ā-ringa (wāhi = tohutohu):
Mo Bark
Bark comes from Suno and takes a different approach from most TTS systems: it is a GPT-style transformer trained as a text-to-audio model rather than a pure text-to-speech one. Because it generates raw audio tokens, it can produce nonverbal sounds — laughing, sighing, crying — as well as background music and sound effects alongside the spoken words. It ships with 100+ speaker presets and handles 13+ languages including English, Chinese, French, German, Hindi, Japanese, and Korean. The trade-off is speed and length: at 350M parameters it runs slowly (~15s per clip) and caps at 200 characters, so it shines for short, emotive, creative audio rather than long narration.
Pai mo: Creative audio content, audiobooks with emotion, sound effects
Ka tirohia katoa Bark ngā oroI te tirohanga
- Ka whakawhanakehia
- Suno
- Ka taea te whakawātea
- MIT
- Karaka
- standard
- Āhuatanga
- slow
- Whakakōrero reo
- Kāore
- reo
- English, Chinese, French, German, Hindi, Italian, Japanese, Korean, Polish, Portuguese, Russian, Spanish, Turkish
- Kāri nui rawa
- 200