PROVIDERLAST 30 DAYS
Gradium voice AI models and benchmarks
Gradium provides real-time speech recognition and synthesis APIs.
- Measured models
- 2
- STTTTS
Overview
Gradium builds and serves both its recognition and synthesis endpoints in-house.
Every model below is measured daily on the same fixed inputs and ranked against the full field, never blended into a company score.
Model lineup
- Speech-to-Text
- Default
- Text-to-Speech
- Default
Speech-to-Text
Full STT dashboardRanked on Time to Final Segment against 24 measured models.
- #15Defaultvia Gradium249 ms
leaders plus Gradium models · 16 other models in the full table
- #26Defaultvia Gradium10.0%
leaders plus Gradium models · 20 other models in the full table
- #20Defaultvia Gradium1979 ms
leaders plus Gradium models · 16 other models in the full table
Text-to-Speech
Full TTS dashboardRanked on Time to First Audio against 30 measured models.
- #15Default384 ms
leaders plus Gradium models · 22 other models in the full table
- #6Default4.6%
leaders plus Gradium models · 23 other models in the full table
Limits of this comparison
Coval measures its default STT and TTS configurations separately.
- The API name `default` is also used by other providers, so Gradium's default configurations are identified by provider and category.
Official resources
- Gradium developer documentation (documentation)
Same datasets, prompts and metric definitions for every model, measured by Coval’s open-source runner. Full methodology on the overview.
Evaluate your own voice agent
Use Coval to test your production configuration, prompts and calls—not only the public benchmark endpoints.