Home / Companies / Gradium / Blog / Post Details
Content Deep Dive

New voices: how we pick the one that wins

Blog post from Gradium

Post Details
Company
Date Published
Author
Gradium
Word Count
2,362
Company Posts That Month
1
Language
English
Hacker News Points
-
Post removed?
No
Summary

Gradium describes voice selection as a structured casting process that combines customer-defined requirements with large-scale generated candidate pools and listener evaluation. Voice targets are specified by locale, demographic characteristics, vocal traits, intended use case, and delivery style, with use case treated as especially influential because listener preferences vary between contexts such as customer support and narration. The company recommends defining the actual listener, creating detailed personas, identifying undesirable brand traits, translating descriptive language into acoustic instructions with native-speaker input, and using realistic evaluation scripts. Flagship voices are selected through diverse generation, crowd-based keeper tests, head-to-head ELO ranking against existing catalog voices, statistical confidence thresholds, and native-speaker checks for accent and pronunciation. The approach emphasizes that voice descriptors and preferences do not transfer reliably across languages or cultures, requiring locale-specific prompts and evaluation, while accent assessment depends on scripts that expose distinguishing sounds and validation by native listeners. Gradium reports a catalog of more than 360 voices across English, French, German, Portuguese, and Spanish, and plans to expand locales, use cases, voice discovery features, and access to its voice-design model.

Trends Found in this Post

No tracked trend matches for this post yet.

Use This Data

Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.