ElevenLabs

ElevenLabs Multilingual V3

ElevenLabs' top-tier TTS — 74 languages, multi-speaker dialogue, emotion tags, audiobook-grade narration.

لا اشتراك
الاعتمادات لا تنتهي أبدا
تعلم المزيد

ادفع مرة واحدة للحصول على أرصدة - استخدمها عبر كل طراز على ZOOOP. · قم بتعبئة الرصيد عندما تحتاج إلى ذلك ، لا حرق شهري.

Powered by ElevenLabs's API on ZOOOP

السمات الرئيسية

74 languages, one model

V3 supports 74 languages — up from ~29 in V2 — covering the vast majority of the world's population. The same voice characteristic carries across languages.

Multi-speaker dialogue

New Text-to-Dialogue API generates natural lifelike dialogue with multiple distinct speakers in a single render — character interactions across languages, with emotional consistency.

Audio tags for direction

Inline tags like [whispering], [sad], [laughs], [shouting] direct the read across languages — a [sad] tag in Spanish lands the same way it does in English.

Hundreds of multilingual voices

Aria, Roger, Sarah, Laura, Charlie, George, Callum, River, Liam, Charlotte, Alice, Matilda, Will, Jessica, Eric, Chris, Brian, Daniel, Lily, Bill — and many more. Each works across all 74 languages.

حالات الاستخدام

Audiobook production

Audiobook production

Long-form narration with audiobook-grade emotional delivery, including subtle tonal shifts across chapters and characters.

Character dialogue

Character dialogue

Multi-speaker Text-to-Dialogue handles full scenes with distinct characters who interact emotionally — useful for animation, games, and audio drama.

Multilingual campaigns

Multilingual campaigns

Generate the same script in 74 languages with consistent voice characteristics. One brand voice, every market, no separate cast per language.

E-learning narration

E-learning narration

Calm explanatory tone with emphasis on key terms — tags let you direct pacing and stress without re-recording.

Podcast intros and ads

Podcast intros and ads

Audiobook-grade fidelity at podcast-ad lengths — drop into existing podcast pipelines without quality drop.

Game character voice

Game character voice

Use audio tags to deliver context-specific reads ([angry], [whispering], [tired]) for in-game lines without a voice cast.

اختيار النموذج الصحيح

Pick the right TTS model for the work. Your credits work everywhere on ZOOOP.

Top quality, 74 languages, multi-speakerElevenLabs V3
Full song with vocals + structureLyria 3 Pro

كيفية استخدام

01

Open ElevenLabs Multilingual V3 from this page or pick it in the Audio Generator.

02

Pick a voice from the library — each works across all 74 languages.

03

Write the script in your target language. Add inline tags like [whispering] or [sad] to direct emotion.

04

Generate. For multi-speaker, switch to Text-to-Dialogue and assign lines per voice.

الغوص العميق

What ElevenLabs Multilingual V3 is good at — and what it's not

ElevenLabs Multilingual V3 is the model that made multilingual TTS production-ready. For most of TTS history, "multilingual" was a checkbox feature — five languages, ten if you were lucky, with the non-English options noticeably stilted. V3 ships with 74 languages — covering the vast majority of the world's population — and the non-English reads hold the same emotional fidelity, pacing, and naturalism as the English ones. Practical effect: a single brand voice now ships across global markets without a separate cast per language and without the off-brand local read that always crept in.

The capability that gets less attention but matters more for production work is audio tags as performance direction. Inline marks like [whispering], [sad], [laughs], [shouting], [angry], [tired] placed directly in the text are read by V3 as directorial instructions and applied across whichever language you're generating in. A [sad] tag in Spanish lands the same way it does in English; a [whispering] instruction in Japanese reads as a hush rather than a quiet baseline. For audiobook narration, character dialogue, and audio drama, this collapses the back-and-forth between "write the line" and "describe how it should sound" — the direction lives in the text itself.

The third flagship capability is the Text-to-Dialogue API. Multi-speaker conversations with distinct characters — each with their own voice — generated as a continuous interaction with emotional consistency. Useful for animation dubs, game cutscenes, audio drama, and any content where the deliverable is character interaction rather than monologue. Pair this with V3's emotion tags and you have a tool that produces what used to require an entire voice cast plus a director.

Voice library is hundreds of multilingual voices — Aria, Roger, Sarah, Laura, Charlie, George, Callum, River, Liam, Charlotte, Alice, Matilda, Will, Jessica, Eric, Chris, Brian, Daniel, Lily, Bill, and many more. Each voice carries its characteristic across all 74 languages, so a deep narrator voice in English stays deep in Mandarin, French, and Korean. For audiobook publishers, e-learning producers, and podcast networks, this is the difference between "AI voice" and "production voice."

Where it's weaker: ultra-low-latency real-time use (live conversational agents under 200ms first-response) is better served by lighter, faster models like Speech-2.8-Turbo from MiniMax. Voice cloning from short samples is supported but specialized models like Chatterbox TTS Multilingual or Index TTS 2 are tuned specifically for that. V3's sweet spot is high-quality narration, multi-speaker dialogue, and multilingual brand work.

A reasonable mental model: V3 is the default for any narration / dialogue work where quality matters more than millisecond latency.

الأسئلة المتداولة

How is V3 different from V2 / Multilingual V2?+

V3 supports 74 languages (up from ~29 in V2), introduces emotion / direction audio tags, ships the Text-to-Dialogue API for multi-speaker scenes, and produces noticeably more natural emotional range. V2 remains a strong baseline; V3 is the upgrade for any new project.

Does V3 work in my language?+

V3 covers 74 languages including English, Chinese (Simplified + Traditional), Japanese, Korean, Spanish, French, German, Portuguese, Hindi, Arabic, Russian, Vietnamese, Thai, Indonesian, Turkish, Polish, Dutch, Norwegian, Danish, and many more — most of the world's commonly used languages.

What are audio tags?+

Inline directorial marks like `[whispering]`, `[laughs]`, `[sad]`, `[angry]`, `[shouting]` placed in the text. V3 reads them as performance direction and applies the emotion across whichever language you're generating in. A [sad] tag in Spanish lands the same way it does in English.

Can V3 do multi-speaker dialogue?+

Yes — the Text-to-Dialogue API generates natural multi-speaker conversations with emotional consistency across speakers and languages. Useful for audio drama, animation dubs, games, and any content with character interactions.

How does V3 compare to other TTS models?+

V3 leads on language coverage (74 languages, more than any competitor) and on direction (audio tags work cross-lingually). For ultra-low latency real-time use, lighter models like Speech-2.8-Turbo are faster. For full audiobook / drama production, V3 is the current quality leader.

المزيد من النماذج

Google
Lyria 3 Pro
Google
OpenAI
GPT Image 2.0
OpenAI
Google
Nano Banana Pro
Google
ByteDance
Seedance V2.0
ByteDance
ByteDance
Seedance V2.0 Fast
ByteDance
ByteDance
Seedance V1.5 Pro
ByteDance
ByteDance
Seedance V1.0 Pro
ByteDance
ByteDance
Seedance V1.0 Pro Fast
ByteDance
ByteDance
Seedance V1.0 Lite
ByteDance
ByteDance
Seedream 5.0 Pro
ByteDance
ByteDance
Seedream 5.0 Lite
ByteDance
ByteDance
Seedream 4.5
ByteDance
ByteDance
Seedream 4
ByteDance
ByteDance
Dreamactor V2
ByteDance
MiniMax
MiniMax H3
MiniMax
MiniMax
Minimax Music V2.6
MiniMax
MiniMax
Minimax Music V2
MiniMax
MiniMax
Speech-2.8-HD
MiniMax
MiniMax
Speech-2.8-Turbo
MiniMax
Kling AI
Kling O3
Kling AI
Kling AI
Kling V3
Kling AI
Kling AI
Kling V3 Pro
Kling AI
Kling AI
Kling V2.6 Pro
Kling AI
Kling AI
Kling V2.6
Kling AI
Kling AI
Kling Lipsync
Kling AI
Kling AI
Kling Avatar V2
Kling AI
Kling AI
Kling O1
Kling AI
Midjourney
Midjourney V8.2
Midjourney
Midjourney
Midjourney V8.1
Midjourney
Midjourney
Midjourney
Midjourney
Alibaba
Happy Horse V1.1
Alibaba
Alibaba
Happy Horse
Alibaba
xAI
Grok Imagine V1.5
xAI
xAI
Grok Imagine
xAI
xAI
xAI TTS
xAI
Google
Veo 3.1
Google
Google
Veo 3.1 Fast
Google
Google
Veo 3
Google
Google
Nano Banana 2
Google
Google
Nano Banana
Google
Google
Lyria 3
Google
Google
Lyria2
Google
Google
Gemini 3.1 Flash TTS
Google
Wan AI
Wan V2.2
Wan AI
Wan AI
Wan V2.2 Turbo
Wan AI
Wan AI
Wan V2.5
Wan AI
Wan AI
Wan V2.6
Wan AI
Wan AI
Wan V2.6 Flash
Wan AI
Wan AI
Wan V2.7
Wan AI
Pixverse AI
Pixverse V6
Pixverse AI
Pixverse AI
Pixverse V5.5
Pixverse AI
Pixverse AI
Pixverse V5
Pixverse AI
Pixverse AI
Pixverse Lipsync
Pixverse AI
Vidu AI
Vidu Q3 Pro
Vidu AI
Vidu AI
Vidu Q3
Vidu AI
Vidu AI
Vidu Q3 Turbo
Vidu AI
Vidu AI
Vidu Q2 Pro
Vidu AI
Vidu AI
Vidu Q2 Turbo
Vidu AI
Luma AI
Luma Ray 2
Luma AI
Luma AI
Luma Ray 2 Flash
Luma AI
Flux AI
Flux 2 Pro
Flux AI
Flux AI
Flux 2
Flux AI
Flux AI
Flux 2 Flash
Flux AI
ElevenLabs
Multilingual V2
ElevenLabs
ElevenLabs
Sound Effects V2
ElevenLabs
Hailuo AI
Hailuo 2.3
Hailuo AI
Hailuo AI
Hailuo 2.3 Fast
Hailuo AI
Hailuo AI
Hailuo 02
Hailuo AI
Lightricks
LTX-2.3 Pro
Lightricks
Lightricks
LTX-2.3 Fast
Lightricks
Lightricks
LTX-2.3
Lightricks
Pika AI
Pika V2.2
Pika AI
Qwen
Qwen3-TTS
Qwen
Inworld
Inworld TTS
Inworld
Bilibili Index
Index TTS 2
Bilibili Index
Resemble AI
Chatterbox TTS Multilingual
Resemble AI
Open Source
ACE-Step
Open Source
Open Source
LUX TTS
Open Source
CassetteAI
Music Generator
CassetteAI
Recraft
Text to Vector V4.1
Recraft
Recraft
Text to Vector V4.1 Pro
Recraft
Recraft
Image to Vector
Recraft