PROVIDERLAST 30 DAYS

official resource

Speechmatics voice AI models and benchmarks

Speechmatics develops multilingual speech recognition services.

Measured models
2
STT
Best STT model#14 / 24
223ms
Default

Overview

The service includes diarization, translation, code switching and vocabulary controls, and can be deployed beyond the public cloud.

Every model below is measured daily on the same fixed inputs and ranked against the full field, never blended into a company score.

Model lineup

Speech-to-Text
Default, Enhanced

Ranked on Time to Final Segment against 24 measured models.

Time to Final Segmentms · lower is better · Speechmatics models markedEvery measured STT model on Time to Final Segment, with Speechmatics's models highlighted.
  1. #1STT RT v564 ms
  2. #3STT 183 ms
  3. #4Nova 399 ms
  4. #5Nova 2101 ms
  5. #6Ink 2108 ms
  6. #14Defaultvia Speechmatics223 ms
  7. #17Enhanced341 ms
leaders plus Speechmatics models · 15 other models in the full table
Word Error Rate% · lower is better · Speechmatics models markedEvery measured STT model on Word Error Rate, with Speechmatics's models highlighted.
  1. #2resonant-13.4%
  2. #3Chirp 34.1%
  3. #4Enhanced4.3%
  4. #6Grok STT4.8%
  5. #7STT 14.8%
  6. #14Defaultvia Speechmatics5.5%
leaders plus Speechmatics models · 20 other models in the full table
Time to First Tokenms · lower is better · Speechmatics models markedEvery measured STT model on Time to First Token, with Speechmatics's models highlighted.
  1. #2Flux1089 ms
  2. #6Nova 31434 ms
  3. #7Nova 21436 ms
  4. #9Defaultvia Speechmatics1473 ms
  5. #12Enhanced1536 ms
leaders plus Speechmatics models · 15 other models in the full table
TTFS vs WER — Speechmatics vs every measured modelEach point is one measured STT model · Speechmatics models highlighted · 30-day averagesTime to Final Segment against Word Error Rate for every measured STT model, with Speechmatics's models highlighted.
Benchmarked models with their Time to Final Segment over the last 30 days.
ModelHostTTFSRank
DefaultSpeechmatics223 ms14th of 24
EnhancedSpeechmatics341 ms17th of 24

Limits of this comparison

Coval measures its default and Enhanced real-time tiers separately.

  • The benchmark does not score translation or every supported language. The model name `default` is also used by other providers, so it is identified together with the Speechmatics provider name.

Official resources

Same datasets, prompts and metric definitions for every model, measured by Coval’s open-source runner. Full methodology on the overview.

Evaluate your own voice agent

Use Coval to test your production configuration, prompts and calls—not only the public benchmark endpoints.

Book a Demo