PROVIDERLAST 30 DAYS
Nari voice AI models and benchmarks
Nari lists 2 STT and TTS models in Coval. Fastest dated mean latency over 30 days: STT: Qwen3 ASR Fast at 46 ms TTFS, with 3.2% WER. TTS: Qwen3 TTS Fast at 70 ms TTFA, with 4.2% WER. Last measured .
Coval measures voice models created or hosted by Nari.
- Measured models
- 2
- STTTTS
Overview
Every model below is measured daily on the same fixed inputs and ranked against the full field, never blended into a company score.
Model lineup
- Speech-to-Text
- serves Qwen3 ASR Fast
- Text-to-Speech
- serves Qwen3 TTS Fast
Speech-to-Text
Full STT dashboardRanked on Time to Final Segment against 27 measured models.
- #1Qwen3 ASR 1.7bDedicated inference. Shared endpoints serve many customers on the same infrastructure, while dedicated endpoints run on hardware reserved for a single customer.39 ms
- #5Qwen3 ASR 1.7bDedicated inference. Shared endpoints serve many customers on the same infrastructure, while dedicated endpoints run on hardware reserved for a single customer.3.9%
- #1Whisper Large v3via BasetenDedicated inference. Shared endpoints serve many customers on the same infrastructure, while dedicated endpoints run on hardware reserved for a single customer.915 ms
- #2Qwen3 ASR 1.7bDedicated inference. Shared endpoints serve many customers on the same infrastructure, while dedicated endpoints run on hardware reserved for a single customer.935 ms
| Model | Host | TTFS | Rank |
|---|---|---|---|
| Qwen3 ASR Fast | Nari | 46 ms | 2nd of 27 |
Text-to-Speech
Full TTS dashboardRanked on Time to First Audio against 28 measured models.
- #4Qwen3 TTS 1.7bDedicated inference. Shared endpoints serve many customers on the same infrastructure, while dedicated endpoints run on hardware reserved for a single customer.105 ms
| Model | Host | TTFA | Rank |
|---|---|---|---|
| Qwen3 TTS Fast | Nari | 70 ms | 2nd of 28 |
How fast are Nari's STT and TTS models?
Qwen3 ASR Fast measures mean 46 ms time to final segment (2nd of 27) among STT systems. Last measured 2026-09-18. Qwen3 TTS Fast measures mean 70 ms time to first audio (2nd of 28) among TTS systems. Last measured 2026-09-18.
How accurate are Nari's STT and TTS models?
Qwen3 ASR Fast measures 3.2% word error rate (2nd of 29) among STT systems. Last measured 2026-09-18. Qwen3 TTS Fast measures 4.2% word error rate (3rd of 28) among TTS systems. Last measured 2026-09-18.
Which Nari model is fastest?
Its fastest dated STT result is Qwen3 ASR Fast at mean 46 ms time to final segment (2nd of 27) among STT systems, with 3.2% WER. Last measured 2026-09-18. Its fastest dated TTS result is Qwen3 TTS Fast at mean 70 ms time to first audio (2nd of 28) among TTS systems, with 4.2% WER. Last measured 2026-09-18.
Same datasets, prompts and metric definitions for every model, measured by Coval’s open-source runner. Full methodology on the overview.
Evaluate your own voice agent
Use Coval to test your production configuration, prompts and calls—not only the public benchmark endpoints.