PROVIDERLAST 30 DAYS

Nari voice AI models and benchmarks

Nari lists 2 STT and TTS models in Coval. Fastest dated mean latency over 30 days: STT: Qwen3 ASR Fast at 46 ms TTFS, with 3.2% WER. TTS: Qwen3 TTS Fast at 70 ms TTFA, with 4.2% WER. Last measured .

Coval measures voice models created or hosted by Nari.

Measured models
2
STTTTS

Overview

Every model below is measured daily on the same fixed inputs and ranked against the full field, never blended into a company score.

Model lineup

Speech-to-Text
serves Qwen3 ASR Fast
Text-to-Speech
serves Qwen3 TTS Fast

Ranked on Time to Final Segment against 27 measured models.

Time to Final Segmentms · lower is better · Nari models markedEvery measured STT model on Time to Final Segment, with Nari's models highlighted.
  1. #1Qwen3 ASR 1.7b39 ms
  2. #3STT RT v557 ms
  3. #4STT 165 ms
  4. #6Nova 388 ms
  5. #7Nova 292 ms
leaders plus Nari models · 20 other models in the full table
Word Error Rate% · lower is better · Nari models markedEvery measured STT model on Word Error Rate, with Nari's models highlighted.
  1. #3resonant-13.4%
  2. #5Qwen3 ASR 1.7b3.9%
  3. #6Chirp 34.1%
leaders plus Nari models · 22 other models in the full table
Time to First Tokenms · lower is better · Nari models markedEvery measured STT model on Time to First Token, with Nari's models highlighted.
  1. #1Whisper Large v3via Baseten915 ms
  2. #2Qwen3 ASR 1.7b935 ms
  3. #4Flux1076 ms
  4. #7Whisper Large v3via Together AI1337 ms
  5. #16Qwen3 ASR Fast1748 ms
leaders plus Nari models · 17 other models in the full table
Benchmarked models with their Time to Final Segment over the last 30 days.
ModelHostTTFSRank
Qwen3 ASR FastNari46 ms2nd of 27

Ranked on Time to First Audio against 28 measured models.

Time to First Audioms · lower is better · Nari models markedEvery measured TTS model on Time to First Audio, with Nari's models highlighted.
  1. #1vui64 ms
  2. #3TTS Flash 291 ms
  3. #4Qwen3 TTS 1.7b105 ms
  4. #6TTS 2181 ms
  5. #7Flash v2.5193 ms
leaders plus Nari models · 21 other models in the full table
Word Error Rate% · lower is better · Nari models markedEvery measured TTS model on Word Error Rate, with Nari's models highlighted.
leaders plus Nari models · 21 other models in the full table
Benchmarked models with their Time to First Audio over the last 30 days.
ModelHostTTFARank
Qwen3 TTS FastNari70 ms2nd of 28

How fast are Nari's STT and TTS models?

Qwen3 ASR Fast measures mean 46 ms time to final segment (2nd of 27) among STT systems. Last measured 2026-09-18. Qwen3 TTS Fast measures mean 70 ms time to first audio (2nd of 28) among TTS systems. Last measured 2026-09-18.

How accurate are Nari's STT and TTS models?

Qwen3 ASR Fast measures 3.2% word error rate (2nd of 29) among STT systems. Last measured 2026-09-18. Qwen3 TTS Fast measures 4.2% word error rate (3rd of 28) among TTS systems. Last measured 2026-09-18.

Which Nari model is fastest?

Its fastest dated STT result is Qwen3 ASR Fast at mean 46 ms time to final segment (2nd of 27) among STT systems, with 3.2% WER. Last measured 2026-09-18. Its fastest dated TTS result is Qwen3 TTS Fast at mean 70 ms time to first audio (2nd of 28) among TTS systems, with 4.2% WER. Last measured 2026-09-18.

Same datasets, prompts and metric definitions for every model, measured by Coval’s open-source runner. Full methodology on the overview.

Evaluate your own voice agent

Use Coval to test your production configuration, prompts and calls—not only the public benchmark endpoints.

Book a Demo