PROVIDERLAST 30 DAYS

official resource

Modulate voice AI models and benchmarks

Modulate develops voice-safety technology and the Velma streaming speech recognition API.

Measured models
1
STT

Overview

The measured endpoint is a first-party multilingual STT service with speaker diarization.

Every model below is measured daily on the same fixed inputs and ranked against the full field, never blended into a company score.

Model lineup

Speech-to-Text
Velma 2 STT Streaming

Ranked on Time to Final Segment against 24 measured models.

Time to Final Segmentms · lower is better · Modulate models markedEvery measured STT model on Time to Final Segment, with Modulate's models highlighted.
  1. #1STT RT v564 ms
  2. #3STT 183 ms
  3. #4Nova 399 ms
  4. #5Nova 2101 ms
  5. #6Ink 2108 ms
leaders plus Modulate models · 16 other models in the full table
Word Error Rate% · lower is better · Modulate models markedEvery measured STT model on Word Error Rate, with Modulate's models highlighted.
leaders plus Modulate models · 20 other models in the full table
Time to First Tokenms · lower is better · Modulate models markedEvery measured STT model on Time to First Token, with Modulate's models highlighted.
  1. #2Flux1089 ms
  2. #6Nova 31434 ms
  3. #7Nova 21436 ms
leaders plus Modulate models · 16 other models in the full table
TTFS vs WER — Modulate vs every measured modelEach point is one measured STT model · Modulate models highlighted · 30-day averagesTime to Final Segment against Word Error Rate for every measured STT model, with Modulate's models highlighted.
Benchmarked models with their Time to Final Segment over the last 30 days.
ModelHostTTFSRank
Velma 2 STT StreamingModulate191 ms11th of 24

Limits of this comparison

Coval limits its provider comparison to Velma's transcription path.

  • Moderation, toxicity detection and other Modulate analysis products are outside Coval's current benchmark metrics.

Official resources

Same datasets, prompts and metric definitions for every model, measured by Coval’s open-source runner. Full methodology on the overview.

Evaluate your own voice agent

Use Coval to test your production configuration, prompts and calls—not only the public benchmark endpoints.

Book a Demo