Home / Companies / Deepgram / Blog / Post Details
Content Deep Dive

Introducing Flux TTS: Conversation-Native Text to Speech for Real Time Voice Agents

Blog post from Deepgram

Post Details
Company
Date Published
Author
Yael Scarlett
Word Count
1,970
Company Posts That Month
20
Language
English
Hacker News Points
-
Post removed?
No
Summary

Deepgram has launched Flux TTS, a generally available text-to-speech model designed for real-time, multi-turn voice agents and offered free through September 12. Unlike narration-oriented TTS systems, Flux TTS is intended to retain conversational context across turns, adapt tone and pacing without SSML or detailed prompting, support interruptions by reporting what callers heard, and allow adjustments to speech characteristics while audio is being generated. Deepgram says the model was trained on conversational speech and uses a high-fidelity neural codec, interleaved text-and-audio generation, and a Mamba state-space architecture to combine expressive delivery, persistent context, and low latency, with first audio reported as low as 80 milliseconds. The company also reports benchmark advantages in word error rates, particularly for difficult production inputs such as account numbers, drug names, dates, currencies, and technical strings. Flux TTS can be deployed through cloud, self-hosted, or on-premises environments and is positioned for regulated industries and noisy settings such as restaurants. It integrates with Deepgram’s Flux STT through a single API configuration, with planned shared state between the speech-to-text and text-to-speech models, while future features include additional languages, voice cloning, emotional controls, and expanded technical documentation.

Trends Found in this Post
Trend Post Mentions Total Month Mentions Posts Companies MoM
Voice AI 20 2,839 275 56 -36%
Real-time 6 4,432 1,050 222 -31%
LLM 2 5,068 1,020 229 -34%
Use This Data

Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.