Product boundary
Whether Deepgram's real-time speech-to-text or text-to-speech APIs fit the required latency, model, language, voice, data, and operating boundaries.
This page evaluates Deepgram's speech-to-text and text-to-speech APIs as two distinct modalities. Flux and Nova recognition paths, Aura and Flux speech-generation paths, add-ons, and Voice Agent APIs have separate capabilities and pricing; support in one modality does not imply support in the other.1234
For: Teams building live transcription or conversational voice systems that will benchmark the exact model, language, latency, and audio path
- The production workload, language, model, provider, or deployment boundary changes
- Current pricing, retention, data use, policy, or regional support changes
- Measured quality, latency, reliability, or operating cost no longer fits