Announcing EVI 3 API: The most customizable speech-to-speech model
Blog post from Hume
EVI 3, the latest empathic voice interface from Hume, offers groundbreaking features in speech-language modeling, enabling expressive speech with any voice without fine-tuning, and introducing hyperrealistic voice cloning with less than 30 seconds of audio. This model surpasses traditional text-to-speech systems by integrating language and voice processing into a single, faster, and higher quality speech-to-speech model, trained on trillions of text tokens and millions of speech hours. EVI 3's interoperability allows seamless integration with other popular large language models like Claude 4 and Gemini 2.5, enhancing its capabilities in AI applications such as assistants and VR characters. The platform supports over 200,000 user-designed voices and offers efficient communication with other systems, ensuring low latency and customizable speech interactions. EVI 3 is accessible at competitive pricing, and developers can explore its features through demos, playgrounds, and comprehensive documentation, setting the stage for innovative AI-driven projects.
| Trend | Post Mentions | Total Month Mentions | Posts | Companies | MoM |
|---|---|---|---|---|---|
| LLM | 16 | 4,922 | 763 | 224 | +11% |
| Voice AI | 5 | 782 | 117 | 41 | -22% |
| AI Model Fine-tuning | 4 | 867 | 189 | 73 | +71% |
| AI Agents | 1 | 2,700 | 582 | 198 | +23% |
| RAG | 1 | 1,131 | 232 | 87 | -9% |
Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.