FreyaTTS

FreyaTTS เสียง

A compact Turkish text-to-speech model that outputs 48 kHz audio without a phonemizer.

ลงทะเบียน สำหรับจำกัดตัวอักษร 5,000 ตัว

หมุนข้อความของคุณในแท็ก SSML เพื่อควบคุมอย่างแม่นยำ:

<speak><prosody rate="slow">Slow speech</prosody></speak>

แท็กที่โมเดลที่เลือกไว้เข้าใจ - คลิกเพื่อวางแท็กลงในข้อความของคุณที่มันเกิดขึ้น:

โมเดลนี้อ่านข้อความธรรมดา ดังนั้น แท็กในบรรทัดจะถูกละเลย สำหรับอารมณ์ที่ใช้แท็ก เปลี่ยนไปใช้โมเดลแสดงออก เช่น Orpheus หรือ Bark

ตั้งค่าการออกเสียงที่กำหนดเอง (คำ = การออกเสียง):

-12 +12
0.5x 2.0x
ใช้ฟรีกับไพเปอร์, VITS, MeloTTS
เสียงที่สร้างขึ้นจะปรากฏที่นี่ เลือกโมเดล พิมพ์ข้อความ และคลิกที่ สร้าง
สร้างเสียงสำเร็จแล้ว
0:00
ดาวน์โหลดเพลง ดาวน์โหลด ลิงก์หมดอายุใน 24 ชั่วโมง
ระดับฟรี: ใช้ส่วนตัว ใบอนุญาตเชิงพาณิชย์จาก $5/เดือน
ทำเสียงนี้เป็นเสียงของตัวเอง โคลนเสียงใน 30 วินาที
รัก TTS.ai บอกเพื่อนๆ

เกี่ยวกับ FreyaTTS

FreyaTTS-small is a 183-million-parameter model built for one language and built well. It is a non-autoregressive conditional flow-matching diffusion transformer that reads Turkish directly at the character level — 92 symbols, no phonemizer and no grapheme-to-phoneme stage, which removes a whole class of mispronunciation that pronunciation dictionaries introduce. It generates in a frozen AudioVAE2 latent space and decodes to 48 kHz mono, more than double the sample rate of the piper Turkish voice, so the output carries treble detail that a 22 kHz model simply cannot represent. On the Freya-TR-Eval benchmark it reaches 8.0% word error rate, placing it ahead of both XTTS-v2 and F5-TTS among open sub-billion-parameter Turkish systems, and it runs fast enough for real-time use at roughly a tenth of real time.

เหมาะสำหรับ: Turkish narration, voice agents, and any Turkish audio that needs high sample-rate output

แสดงทั้งหมด FreyaTTS เสียง

เพียงแค่มองดู

ผู้พัฒนา
Freya
ใบอนุญาต
Apache 2.0
สัตว์
free
ความเร็ว
fast
เสียง
ไม่มี
ภาษา
Turkish
จำนวนตัวอักษรสูงสุด
2000

FreyaTTS เสียง

Freya (Turkish)

Turkish
ค่ามาตรฐาน Female

FreyaTTS คำถามที่พบบ่อย

It was trained from scratch on Turkish speech alone rather than adapted from a multilingual model. Specialising lets a 183M-parameter model compete with far larger multilingual systems on Turkish, but it means the model has no ability to read other languages — requests in another language are rejected rather than mispronounced.

Most open TTS models emit 22.05 kHz or 24 kHz, which caps reproducible audio at around 11-12 kHz and audibly dulls sibilants. FreyaTTS decodes to 48 kHz, the standard sample rate for video and broadcast, so its output drops into a production timeline without upsampling.

No. FreyaTTS has a single fixed speaker and does not support cloning. For a cloned Turkish voice use one of our zero-shot cloning models instead.

Many TTS systems first convert text into phonemes using a pronunciation dictionary, and anything missing from that dictionary — new words, names, loanwords — gets guessed. FreyaTTS reads the characters themselves, so Turkish spelling, which is highly regular, maps to sound without that lossy middle step.
← เสียงทั้งหมด