PROVIDERLAST 30 DAYS
Mistral AI voice AI models and benchmarks
Mistral AI lists 1 STT model in Coval. Its fastest dated STT result is Voxtral Mini Transcribe Realtime 2602 at mean 401 ms time to final segment (22nd of 28) among STT systems, with 6.2% WER. Results cover the last 30 days. Last measured .
Mistral AI develops the Voxtral audio model family and hosts the measured real-time transcription endpoint in Europe.
- Measured models
- 1
- STT
Overview
The full dated Voxtral model identifier remains visible, allowing later releases to coexist instead of silently replacing earlier measurements.
Every model below is measured daily on the same fixed inputs and ranked against the full field, never blended into a company score.
Model lineup
- Speech-to-Text
- Voxtral Mini Transcribe Realtime 2602
Speech-to-Text
Full STT dashboardRanked on Time to Final Segment against 28 measured models.
- #1Qwen3 ASR 1.7bDedicated inference. Shared endpoints serve many customers on the same infrastructure, while dedicated endpoints run on hardware reserved for a single customer.39 ms
- #6Qwen3 ASR 1.7bDedicated inference. Shared endpoints serve many customers on the same infrastructure, while dedicated endpoints run on hardware reserved for a single customer.4.1%
- #1Whisper Large v3via BasetenDedicated inference. Shared endpoints serve many customers on the same infrastructure, while dedicated endpoints run on hardware reserved for a single customer.912 ms
- #2Qwen3 ASR 1.7bDedicated inference. Shared endpoints serve many customers on the same infrastructure, while dedicated endpoints run on hardware reserved for a single customer.931 ms
| Model | Host | TTFS | Rank |
|---|---|---|---|
| Voxtral Mini Transcribe Realtime 2602 | Mistral | 401 ms | 22nd of 28 |
How fast are Mistral AI's STT models?
Voxtral Mini Transcribe Realtime 2602 measures mean 401 ms time to final segment (22nd of 28) among STT systems. Last measured 2026-09-15.
How accurate are Mistral AI's STT models?
Voxtral Mini Transcribe Realtime 2602 measures 6.2% word error rate (21st of 30) among STT systems. Last measured 2026-09-15.
Which Mistral AI model is fastest?
Its fastest dated STT result is Voxtral Mini Transcribe Realtime 2602 at mean 401 ms time to final segment (22nd of 28) among STT systems, with 6.2% WER. Last measured 2026-09-15.
Limits of this comparison
- The benchmark covers one Voxtral STT release, not every Voxtral or Mistral model.
Official resources
Same datasets, prompts and metric definitions for every model, measured by Coval’s open-source runner. Full methodology on the overview.
Evaluate your own voice agent
Use Coval to test your production configuration, prompts and calls—not only the public benchmark endpoints.