PROVIDERLAST 30 DAYS

official resource

Mistral AI voice AI models and benchmarks

Mistral AI develops the Voxtral audio model family and hosts the measured real-time transcription endpoint in Europe.

Overview

The full dated Voxtral model identifier remains visible, allowing later releases to coexist instead of silently replacing earlier measurements.

Every model below is measured daily on the same fixed inputs and ranked against the full field, never blended into a company score.

Model lineup

Ranked on Time to Final Segment against 24 measured models.

Time to Final Segmentms · lower is better · Mistral AI models markedEvery measured STT model on Time to Final Segment, with Mistral AI's models highlighted.
leaders plus Mistral AI models · 16 other models in the full table
Word Error Rate% · lower is better · Mistral AI models markedEvery measured STT model on Word Error Rate, with Mistral AI's models highlighted.
leaders plus Mistral AI models · 20 other models in the full table
Time to First Tokenms · lower is better · Mistral AI models markedEvery measured STT model on Time to First Token, with Mistral AI's models highlighted.
leaders plus Mistral AI models · 16 other models in the full table
TTFS vs WER — Mistral AI vs every measured modelEach point is one measured STT model · Mistral AI models highlighted · 30-day averagesTime to Final Segment against Word Error Rate for every measured STT model, with Mistral AI's models highlighted.
Benchmarked models with their Time to Final Segment over the last 30 days.
ModelHostTTFSRank
Voxtral Mini Transcribe Realtime 2602Mistral356 ms18th of 24

Limits of this comparison

  • The benchmark covers one Voxtral STT release, not every Voxtral or Mistral model.

Official resources

Same datasets, prompts and metric definitions for every model, measured by Coval’s open-source runner. Full methodology on the overview.

Evaluate your own voice agent

Use Coval to test your production configuration, prompts and calls—not only the public benchmark endpoints.

Book a Demo