Deepgram launches Flux TTS and passes $100M ARR
On 2026-08-12 Deepgram launched Flux TTS, a "conversation-native" text-to-speech model for voice agents that keeps context and voice consistency across turns. It responds in as little as 80 ms and reports exactly what the user heard on interruption. It completes the Flux line after Flux STT (Oct 2025, billed as the first conversational speech recognition model) and Flux Multilingual (Apr 2026). Deepgram said it had passed $100M in annual recurring revenue.
Key facts
- Endpoint /v2/speak (WebSocket + REST); voices flux-{voice}-en, 39 English voices
- $0.045 per 1K chars PAYG after a free period ending 2026-09-12
- Self-hosted GA 2026-08-26 with speed and expressivity controls
- Flux STT: flux-general-en ($0.0065/min) and flux-general-multi (10 languages, $0.0078/min)
- Deepgram passed $100M ARR
What happened
Deepgram extended the turn-aware Flux design from speech recognition to speech synthesis. That gives it a full in-house agent stack (Flux STT + LLM + Flux TTS) behind its Voice Agent API.
Why it matters
Voice-agent vendors are building TTS around dialogue state (turns, interruptions, what was actually heard) rather than isolated sentences. Deepgram's $100M ARR also shows the market for speech APIs is growing.
Changelog
- 2026-09-29: created
Models
- Deepgram Flux TTS Deepgram · current
Related events
- Inworld Realtime TTS-2 reaches GA with audio-aware, prompt-directed speech ★★★
- Cartesia Sonic-3.6 goes GA and tops the Artificial Analysis Speech Arena ★★★
- Speechmatics launches Agent STT, powered by its Linden model, for voice agents ★★
Sources (4)
- officialDeepgram: Text-to-Speech comes of age (Flux TTS launch)
- docsDeepgram docs: Flux TTS overview
- officialDeepgram: Flux Multilingual launch (2026-04-29)
- officialDeepgram pricing
id: 2026-08-12-deepgram-flux-tts · updated 2026-09-29 · open in the interactive timeline