PROVIDERLAST 30 DAYS

official resource

AssemblyAI voice AI models and benchmarks

AssemblyAI builds managed speech recognition APIs.

Measured models
2
STT

Overview

AssemblyAI builds and serves its own endpoints, with multilingual, diarization and keyterm capabilities.

Every model below is measured daily on the same fixed inputs and ranked against the full field, never blended into a company score.

Model lineup

Ranked on Time to Final Segment against 24 measured models.

Time to Final Segmentms · lower is better · AssemblyAI models markedEvery measured STT model on Time to Final Segment, with AssemblyAI's models highlighted.
  1. #1STT RT v564 ms
  2. #3STT 183 ms
  3. #4Nova 399 ms
  4. #5Nova 2101 ms
  5. #6Ink 2108 ms
leaders plus AssemblyAI models · 16 other models in the full table
Word Error Rate% · lower is better · AssemblyAI models markedEvery measured STT model on Word Error Rate, with AssemblyAI's models highlighted.
leaders plus AssemblyAI models · 20 other models in the full table
Time to First Tokenms · lower is better · AssemblyAI models markedEvery measured STT model on Time to First Token, with AssemblyAI's models highlighted.
  1. #2Flux1089 ms
  2. #6Nova 31434 ms
  3. #7Nova 21436 ms
leaders plus AssemblyAI models · 16 other models in the full table
TTFS vs WER — AssemblyAI vs every measured modelEach point is one measured STT model · AssemblyAI models highlighted · 30-day averagesTime to Final Segment against Word Error Rate for every measured STT model, with AssemblyAI's models highlighted.
Benchmarked models with their Time to Final Segment over the last 30 days.
ModelHostTTFSRank
Universal StreamingAssemblyAI
Universal 3.5 ProAssemblyAI146 ms8th of 24

Limits of this comparison

Coval tracks its accuracy-oriented Universal model and its Universal Streaming endpoint separately so a product mode does not become a single provider score.

  • Coval's measurements do not cover every AssemblyAI audio-intelligence feature, language or asynchronous workflow.

Official resources

Same datasets, prompts and metric definitions for every model, measured by Coval’s open-source runner. Full methodology on the overview.

Evaluate your own voice agent

Use Coval to test your production configuration, prompts and calls—not only the public benchmark endpoints.

Book a Demo