PROVIDERLAST 30 DAYS
Deepgram voice AI models and benchmarks
Deepgram develops and hosts speech recognition and synthesis APIs.
- Measured models
- 5
- STTTTS
Overview
Deepgram's lineup spans the Nova and Flux recognition families and Aura synthesis, with English, multilingual and voice variants offered as distinct models.
Every model below is measured daily on the same fixed inputs and ranked against the full field, never blended into a company score.
Model lineup
- Speech-to-Text
- Nova 2, Nova 3, Flux, Flux Multilingual
- Text-to-Speech
- Aura 2
Speech-to-Text
Full STT dashboardRanked on Time to Final Segment against 24 measured models.
Text-to-Speech
Full TTS dashboardRanked on Time to First Audio against 30 measured models.
- #6Default4.6%
Limits of this comparison
Coval measures Nova and Flux transcription models and Aura speech output.
- Coval does not evaluate every Deepgram domain model, language, voice or endpointing configuration.
Official resources
- Deepgram model documentation (documentation)
Same datasets, prompts and metric definitions for every model, measured by Coval’s open-source runner. Full methodology on the overview.
Evaluate your own voice agent
Use Coval to test your production configuration, prompts and calls—not only the public benchmark endpoints.