Introducing Sonic-3.6
Blog post from Cartesia
Cartesia announced Sonic-3.6, a multilingual text-to-speech model that it says improves substantially on Sonic-3.5 in naturalness, voice quality, emotional delivery, and context-aware intonation. The company reports that listeners preferred the model in up to 93% of blind comparisons across 15 locales, including a 92% preference over Eleven v3 for U.S. English, and says it ranks first on Artificial Analysis leaderboards. Sonic-3.6 supports 44 languages, 61 locales, more than 500 preset voices, and newly added Odia and Urdu, while offering features such as Hindi-English code-switching, locale-specific pronunciation of dates and numbers, and improved accent retention in instant voice clones. Designed for enterprise deployment, it reportedly responds in under 90 milliseconds, generates speech nearly twice as fast as v3 Conversational, supports cloud, on-premises, and localized deployments, and includes a 99.9% uptime SLA and compliance coverage for SOC 2, PCI, HIPAA, and GDPR. The model is generally available through Cartesia’s playground and API.
No tracked trend matches for this post yet.
Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.