Dia TTS TTS
A 1.6B-parameter model purpose-built for generating natural multi-speaker dialogue, not just single-voice narration.
உங்களின் உரை SSML ஒட்டுகளில் சரியான கட்டுப்பாட்டிற்காக மடி:
<speak><prosody rate="slow">Slow speech</prosody></speak>
தேர்ந்தெடுக்கப்பட்ட மாதிரி புரிந்து கொள்ளும் குறிகள் - உங்கள் உரைக்குள் ஒன்றை இழுக்க க்ளிக் செய்யவும் அது நடக்கும் இடம்:
இந்த மாதிரி வெறும் உரை படிக்கிறது, எனவே உள்ளமைந்த குறிகள் புறக்கணிக்கப்படும். குறி அடிப்படை உணர்வுகளுக்கு, Orpheus அல்லது Bark போன்ற ஒரு வெளிப்படுத்தும் மாதிரிக்கு மாறவும்.
தனிப்பயன் உச்சரிப்புகளை வரையறுக்கவும் (வார்த்தை = உச்சரிப்பு):
& பற்றி Dia TTS
Dia by Nari Labs is a 1.6-billion-parameter text-to-speech model designed from the ground up for dialogue rather than monologue. It generates conversations between two speakers with realistic turn-taking, prosody, and emotional expression, producing audio that sounds like a real exchange instead of two voices read separately. Architecturally it pairs an autoregressive transformer with the Descript Audio Codec (DAC) for waveform generation. Dia is a strong fit for podcast-style content, scripted audiobook dialogue, and conversational scenes, and is released under Apache 2.0. Generations are heavier than single-voice models, so it favors quality over raw speed.
சிறந்த: Podcasts, audiobook dialogues, conversational content
அனைத்தையும் உலாவுக Dia TTS குரல்கள்ஒரு பார்வை
- உருவாக்குநர்
- Nari Labs
- உரிமம்
- Apache 2.0
- மிருகம்
- standard
- வேகம்
- medium
- குரல் ஒப்புமை
- இல்லை
- மொழிகள்
- English
- அதிகபட்ச எழுத்துக்கள்
- 800