Home / Companies / Hume / Blog / Post Details
Content Deep Dive

Announcing EVI 3 API: The most customizable speech-to-speech model

Blog post from Hume

Post Details
Company
Date Published
Author
Hume AI Team
Word Count
993
Company Posts That Month
5
Language
English
Hacker News Points
-
Post removed?
No
Summary

EVI 3, the latest empathic voice interface from Hume, offers groundbreaking features in speech-language modeling, enabling expressive speech with any voice without fine-tuning, and introducing hyperrealistic voice cloning with less than 30 seconds of audio. This model surpasses traditional text-to-speech systems by integrating language and voice processing into a single, faster, and higher quality speech-to-speech model, trained on trillions of text tokens and millions of speech hours. EVI 3's interoperability allows seamless integration with other popular large language models like Claude 4 and Gemini 2.5, enhancing its capabilities in AI applications such as assistants and VR characters. The platform supports over 200,000 user-designed voices and offers efficient communication with other systems, ensuring low latency and customizable speech interactions. EVI 3 is accessible at competitive pricing, and developers can explore its features through demos, playgrounds, and comprehensive documentation, setting the stage for innovative AI-driven projects.

Trends Found in this Post
Trend Post Mentions Total Month Mentions Posts Companies MoM
LLM 16 4,922 763 224 +11%
Voice AI 5 782 117 41 -22%
AI Model Fine-tuning 4 867 189 73 +71%
AI Agents 1 2,700 582 198 +23%
RAG 1 1,131 232 87 -9%
Use This Data

Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.