Parler TTS

Parler TTS เสียง

Describe the voice you want in plain English and Parler generates speech matching that description.

ลงทะเบียน สำหรับจำกัดตัวอักษร 5,000 ตัว

หมุนข้อความของคุณในแท็ก SSML เพื่อควบคุมอย่างแม่นยำ:

<speak><prosody rate="slow">Slow speech</prosody></speak>

แท็กที่โมเดลที่เลือกไว้เข้าใจ - คลิกเพื่อวางแท็กลงในข้อความของคุณที่มันเกิดขึ้น:

โมเดลนี้อ่านข้อความธรรมดา ดังนั้น แท็กในบรรทัดจะถูกละเลย สำหรับอารมณ์ที่ใช้แท็ก เปลี่ยนไปใช้โมเดลแสดงออก เช่น Orpheus หรือ Bark

ตั้งค่าการออกเสียงที่กำหนดเอง (คำ = การออกเสียง):

-12 +12
0.5x 2.0x
ใช้ฟรีกับไพเปอร์, VITS, MeloTTS
เสียงที่สร้างขึ้นจะปรากฏที่นี่ เลือกโมเดล พิมพ์ข้อความ และคลิกที่ สร้าง
สร้างเสียงสำเร็จแล้ว
0:00
ดาวน์โหลดเพลง ดาวน์โหลด ลิงก์หมดอายุใน 24 ชั่วโมง
ระดับฟรี: ใช้ส่วนตัว ใบอนุญาตเชิงพาณิชย์จาก $5/เดือน
ทำเสียงนี้เป็นเสียงของตัวเอง โคลนเสียงใน 30 วินาที
รัก TTS.ai บอกเพื่อนๆ

เกี่ยวกับ Parler TTS

Parler TTS, developed by Hugging Face, replaces voice presets with natural-language control: instead of picking from a fixed list, you write a description such as "a warm female voice with a slight British accent, speaking slowly and clearly," and the model synthesizes speech to match. This makes it unusually flexible for creative work where you need a specific, custom voice character without recording or cloning anyone. It is an 880M-parameter transformer encoder-decoder trained on roughly 45,000 hours of speech, and it is released under the permissive Apache 2.0 license. Parler is English-focused and best suited to applications that benefit from on-demand, describable voice characteristics.

เหมาะสำหรับ: Creative applications where you need custom voice characteristics

แสดงทั้งหมด Parler TTS เสียง

เพียงแค่มองดู

ผู้พัฒนา
Hugging Face
ใบอนุญาต
Apache 2.0
สัตว์
standard
ความเร็ว
medium
เสียง
ไม่มี
ภาษา
English
จำนวนตัวอักษรสูงสุด
500

Parler TTS เสียง

Default

English
ค่ามาตรฐาน Neutral

Parler TTS คำถามที่พบบ่อย

You describe it in natural language — gender, accent, pace, tone, and recording quality — and Parler generates speech matching the description. There are no preset voices to choose from.

Hugging Face. It is an 880M-parameter transformer encoder-decoder trained on around 45,000 hours of speech and released under Apache 2.0.

No. Parler generates a voice from a text description rather than from a reference recording. For cloning a specific voice, use a model like Chatterbox or GPT-SoVITS.
← เสียงทั้งหมด