ElevenLabs Audio Tags: More control over AI Voices
Blog post from ElevenLabs
ElevenLabs has released Eleven v3, an alpha research preview of a new AI voice model that introduces Audio Tags to enhance control over emotion, pacing, and sound effects in text-to-speech applications. These tags, which are words enclosed in square brackets, allow users to direct the AI voice to express emotions, delivery styles, and even nonverbal cues like pauses and tone, thus elevating the expressiveness of generated speech. This feature is particularly useful for producing immersive audiobooks, interactive characters, and dialogue-driven media, offering precise control over audio delivery. Despite Professional Voice Clones (PVCs) not being fully optimized for Eleven v3, Instant Voice Clones (IVCs) or designed voices can be utilized to explore v3's features. Available in the ElevenLabs UI and through a public API, Eleven v3 is currently offered at a discounted rate, encouraging experimentation with its enhanced capabilities.
| Trend | Post Mentions | Total Month Mentions | Posts | Companies | MoM |
|---|---|---|---|---|---|
| Real-time | 1 | 4,075 | 1,042 | 211 | +22% |
| Vector Search | 1 | 1,525 | 253 | 110 | -6% |
Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.