Tag kang dipahami model kang dipilih - klik kanggo ngethok siji ing teks sampeyan ing ngendi iku kedadeyan:
Model iki maca teks biasa, mula tag ing baris diabaikan. Kanggo emosi berbasis tag, ganti menyang model ekspresif kaya Orpheus utawa Bark.
Cues pangiriman basa alami. Dihormati déning Qwen3-TTS saiki; modél ekspresif liyane bakal teka.
Model iki ora nyokong instruksi gaya — ganti menyang model sing nyokong (kayata Qwen3-TTS).
Nyathet tembung-tembung standar (kata = tembung):
-12
+12
Kontrol ekspresifitas (0 = netral, 1 = banget ekspresif)
Klasifikasi-gratis bobot guide (luwih dhuwur = luwih prompt-following)
Nyathet gaya swara ing basa alami (Parler nggunakake iki tinimbang swara kang wis ditetepake)
Dialog Format: Gunake tag [S1] lan [S2] kanggo nyathet panyatur kang béda. Conto: [S1] Halo! [S2] Halo, apa kabarmu?
Iki ya iku model swara premium, kang kasedhiya ing rencana bayar apa wae. Sampeyan bisa uga ndeleng prabédan swara kanthi gratis nganggo tombol main ing sisihé pemilih swara.
FreyaTTS-small is a 183-million-parameter model built for one language and built well. It is a non-autoregressive conditional flow-matching diffusion transformer that reads Turkish directly at the character level — 92 symbols, no phonemizer and no grapheme-to-phoneme stage, which removes a whole class of mispronunciation that pronunciation dictionaries introduce. It generates in a frozen AudioVAE2 latent space and decodes to 48 kHz mono, more than double the sample rate of the piper Turkish voice, so the output carries treble detail that a 22 kHz model simply cannot represent. On the Freya-TR-Eval benchmark it reaches 8.0% word error rate, placing it ahead of both XTTS-v2 and F5-TTS among open sub-billion-parameter Turkish systems, and it runs fast enough for real-time use at roughly a tenth of real time.
Paling apik kanggo: Turkish narration, voice agents, and any Turkish audio that needs high sample-rate output
It was trained from scratch on Turkish speech alone rather than adapted from a multilingual model. Specialising lets a 183M-parameter model compete with far larger multilingual systems on Turkish, but it means the model has no ability to read other languages — requests in another language are rejected rather than mispronounced.
Most open TTS models emit 22.05 kHz or 24 kHz, which caps reproducible audio at around 11-12 kHz and audibly dulls sibilants. FreyaTTS decodes to 48 kHz, the standard sample rate for video and broadcast, so its output drops into a production timeline without upsampling.
No. FreyaTTS has a single fixed speaker and does not support cloning. For a cloned Turkish voice use one of our zero-shot cloning models instead.
Many TTS systems first convert text into phonemes using a pronunciation dictionary, and anything missing from that dictionary — new words, names, loanwords — gets guessed. FreyaTTS reads the characters themselves, so Turkish spelling, which is highly regular, maps to sound without that lossy middle step.
Ing taun 2000, dhèwèké diundang kanggo main ing filem The 1500 Days of Christmas.
Ora perlu sandi — kita bakal ngirim link menyang email kanggo nyetel sandi sabanjuré.
Batas Lapisan Bebas Ditempuh
Sampeyan wis nggunakake 5,000 karakter gratis saben dina. Cipta akun gratis kanggo entuk 15,000 karakter bonus ditambah 10,000 karakter gratis saben wulan.