PROVIDERLAST 30 DAYS

official resource

Together AI voice AI models and benchmarks

Together AI hosts the measured endpoints for open-weight speech recognition models created by NVIDIA and OpenAI.

Measured models
3
STT

Overview

Model accuracy depends on the weights and inference configuration, while latency also reflects Together's serving runtime, region and capacity.

Every model below is measured daily on the same fixed inputs and ranked against the full field, never blended into a company score.

Ranked on Time to Final Segment against 24 measured models.

Time to Final Segmentms · lower is better · Together AI models markedEvery measured STT model on Time to Final Segment, with Together AI's models highlighted.
  1. #1STT RT v564 ms
  2. #3STT 183 ms
  3. #4Nova 399 ms
  4. #5Nova 2101 ms
  5. #6Ink 2108 ms
leaders plus Together AI models · 16 other models in the full table
Word Error Rate% · lower is better · Together AI models markedEvery measured STT model on Word Error Rate, with Together AI's models highlighted.
leaders plus Together AI models · 18 other models in the full table
Time to First Tokenms · lower is better · Together AI models markedEvery measured STT model on Time to First Token, with Together AI's models highlighted.
leaders plus Together AI models · 16 other models in the full table
TTFS vs WER — Together AI vs every measured modelEach point is one measured STT model · Together AI models highlighted · 30-day averagesTime to Final Segment against Word Error Rate for every measured STT model, with Together AI's models highlighted.
Benchmarked models with their Time to Final Segment over the last 30 days.
ModelHostTTFSRank
Nemotron 3.5 ASR StreamingTogether AI
Parakeet TDT 0.6B v3Together AI70 ms2nd of 24
Whisper Large v3Together AI180 ms10th of 24

Limits of this comparison

  • The displayed latency cannot be generalized to self-hosting, NVIDIA NIM or another provider serving the same model.

Official resources

Same datasets, prompts and metric definitions for every model, measured by Coval’s open-source runner. Full methodology on the overview.

Evaluate your own voice agent

Use Coval to test your production configuration, prompts and calls—not only the public benchmark endpoints.

Book a Demo