FreyaTTS TTS
A compact Turkish text-to-speech model that outputs 48 kHz audio without a phonemizer.
Wrap uw tekst in SSML-tags voor nauwkeurige controle:
<speak><prosody rate="slow">Slow speech</prosody></speak>
Tags het geselecteerde model begrijpt
Dit model leest platte tekst, dus inline tags worden genegeerd. Voor emotie op basis van tags, schakel naar een expressief model zoals Orpheus of Bark.
Definieer aangepaste uitspraaken (woord = uitspraak):
Info FreyaTTS
FreyaTTS-small is a 183-million-parameter model built for one language and built well. It is a non-autoregressive conditional flow-matching diffusion transformer that reads Turkish directly at the character level — 92 symbols, no phonemizer and no grapheme-to-phoneme stage, which removes a whole class of mispronunciation that pronunciation dictionaries introduce. It generates in a frozen AudioVAE2 latent space and decodes to 48 kHz mono, more than double the sample rate of the piper Turkish voice, so the output carries treble detail that a 22 kHz model simply cannot represent. On the Freya-TR-Eval benchmark it reaches 8.0% word error rate, placing it ahead of both XTTS-v2 and F5-TTS among open sub-billion-parameter Turkish systems, and it runs fast enough for real-time use at roughly a tenth of real time.
Beste voor: Turkish narration, voice agents, and any Turkish audio that needs high sample-rate output
Alles doorbladeren FreyaTTS stemmenIn een oogopslag
- Ontwikkelaar
- Freya
- Licentie
- Apache 2.0
- Niveau
- free
- Snelheid
- fast
- Klonen van stemmen
- Nee
- Talen
- Turkish
- Max. tekens
- 2000