PROVIDERLAST 30 DAYS
Mistral AI voice AI models and benchmarks
Mistral AI develops the Voxtral audio model family and hosts the measured real-time transcription endpoint in Europe.
- Measured models
- 1
- STT
Overview
The full dated Voxtral model identifier remains visible, allowing later releases to coexist instead of silently replacing earlier measurements.
Every model below is measured daily on the same fixed inputs and ranked against the full field, never blended into a company score.
Model lineup
- Speech-to-Text
- Voxtral Mini Transcribe Realtime 2602
Speech-to-Text
Full STT dashboardRanked on Time to Final Segment against 24 measured models.
leaders plus Mistral AI models · 16 other models in the full table
leaders plus Mistral AI models · 20 other models in the full table
leaders plus Mistral AI models · 16 other models in the full table
| Model | Host | TTFS | Rank |
|---|---|---|---|
| Voxtral Mini Transcribe Realtime 2602 | Mistral | 356 ms | 18th of 24 |
Limits of this comparison
- The benchmark covers one Voxtral STT release, not every Voxtral or Mistral model.
Official resources
- Mistral audio transcription (documentation)
Same datasets, prompts and metric definitions for every model, measured by Coval’s open-source runner. Full methodology on the overview.
Evaluate your own voice agent
Use Coval to test your production configuration, prompts and calls—not only the public benchmark endpoints.