Inline Audio Tags as director notes
Write [whispers], [excited], [laughs], [sighs], [sarcastically] in copy. For narration and ad reads that need specific emotion, less “generate again and hope.”
ElevenLabs GA’s most expressive TTS: inline Audio Tags, multi-speaker dialogue, 70+ languages; 4.9% error rate. Default for narration, dialogue, and performance reads.

GA vs Alpha: 72% user preference; digit, symbol, and notation error rate 15.3%→4.9% — fewer retakes when voiceover misreads amounts and dates.

Inline Audio Tags for emotion, multi-speaker dialogue, 70+ languages in one workflow — when narration feels flat, characters talk over each other, or accents drift across languages, fix tags and speakers first instead of blind rerolls.
Write [whispers], [excited], [laughs], [sighs], [sarcastically] in copy. For narration and ad reads that need specific emotion, less “generate again and hope.”
Text to Dialogue handles multi-character turns, interruptions, and emotional handoffs. Podcasts, game cutscenes, story demos — fewer “one voice playing two roles.”
Global content without swapping models per locale. Localized VO and multilingual ads: align specs, then edit copy and tags — not reopen a whole pipeline.
Narration to dialogue to ad reads — different deliverables, same failure cost: rerecords and missed deadlines.
Use tags for breath and emotional arc when needed. Flat machine read or weak climax breaks immersion; ElevenLabs v3 defaults to narrative expressiveness.
Multi-speaker shared context. Characters talking over each other or wrong emotion kills cutscenes; dialogue mode beats stitching single-speaker TTS.
Name excitement, restraint, irony with tags. Wrong read tone means rerecording the whole spot; set tags first, listen, less blind sampling.
Go to Audio and select ElevenLabs v3.
Body copy + lowercase English bracket tags, e.g. [whispers] Something’s coming… [sighs]. Multi-character: use dialogue flow to split speakers.
Listen for emotion and diction first; if off, change tags/wording and regenerate — don’t rewrite the whole script blindly.
Align before you start: free credits, model name, tag syntax, when to default to it.
Audio Tags · multi-speaker · 70+ languages — alongside image/video models, open and go.
Use ElevenLabs v3 freeNo ElevenLabs API key required.