Newly-released Google's Gemini 3.8 Flash TTS tops Hume's Real-World VoiceEQ leaderboard
Blog post from Hume
Hume’s Real-World VoiceEQ benchmark evaluates text-to-speech systems on practical qualities beyond single-sentence audio, including long-form stability, expressivity, naturalness, speaker identity, multilingual performance, controllability, and voice replication. Its new Frontier Metric combines expressivity and reliability, reflecting a trade-off in which highly expressive models often sacrifice consistency; under this measure, Google’s Gemini 3.8 Flash leads by pairing top expressivity with strong reliability. Gemini 3.8 Flash and Flash-Lite rank first and second overall among 34 models, with particularly large improvements in extended long-form stability, while Flash also leads in voice design and style-tag adherence. Flash-Lite’s voice replication performance is closer to the field average, and Flash has relative weaknesses in producing younger-sounding voices and controlling volume. Hume argues that public benchmarks provide broad comparisons, while private, held-out evaluations through its Kairos platform help research teams diagnose specific model failures, tailor tests to real use cases, and measure improvement without allowing models to train on benchmark materials.
| Trend | Post Mentions | Total Month Mentions | Posts | Companies | MoM |
|---|---|---|---|---|---|
| Gemini 3.8 Flash | 6 | No monthly metrics for this publish month. | |||
| Gemini 3.8 Flash TTS | 1 | No monthly metrics for this publish month. | |||
| Voice AI | 1 | 324 | 41 | 16 | -89% |
Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.