PROVIDERLAST 30 DAYS
Smallest.ai voice AI models and benchmarks
Smallest.ai (Smallest, smallest.ai) lists 2 STT and TTS models in Coval. Fastest dated mean latency over 30 days: STT: Pulse at 212 ms TTFS, with 5.3% WER. TTS: Lightning v3.1 Pro at 362 ms TTFA, with 4.4% WER. Last measured .
Smallest.ai develops real-time speech recognition and synthesis.
- Measured models
- 2
- STTTTS
Overview
Smallest.ai's Pulse recognizer and Lightning synthesizer are both built and served in-house.
Every model below is measured daily on the same fixed inputs and ranked against the full field, never blended into a company score.
Model lineup
- Speech-to-Text
- Pulse
- Text-to-Speech
- Lightning v3.1 Pro
Speech-to-Text
Full STT dashboardRanked on Time to Final Segment against 28 measured models.
- #1Qwen3 ASR 1.7bDedicated inference. Shared endpoints serve many customers on the same infrastructure, while dedicated endpoints run on hardware reserved for a single customer.39 ms
- #6Qwen3 ASR 1.7bDedicated inference. Shared endpoints serve many customers on the same infrastructure, while dedicated endpoints run on hardware reserved for a single customer.4.1%
- #1Whisper Large v3via BasetenDedicated inference. Shared endpoints serve many customers on the same infrastructure, while dedicated endpoints run on hardware reserved for a single customer.912 ms
- #2Qwen3 ASR 1.7bDedicated inference. Shared endpoints serve many customers on the same infrastructure, while dedicated endpoints run on hardware reserved for a single customer.931 ms
Text-to-Speech
Full TTS dashboardRanked on Time to First Audio against 28 measured models.
- #4Qwen3 TTS 1.7bDedicated inference. Shared endpoints serve many customers on the same infrastructure, while dedicated endpoints run on hardware reserved for a single customer.106 ms
| Model | Host | TTFA | Rank |
|---|---|---|---|
| Lightning v3.1 Pro | Smallest | 362 ms | 18th of 28 |
How fast are Smallest.ai's STT and TTS models?
Pulse measures mean 212 ms time to final segment (16th of 28) among STT systems. Last measured 2026-09-15. Lightning v3.1 Pro measures mean 362 ms time to first audio (18th of 28) among TTS systems. Last measured 2026-09-15.
How accurate are Smallest.ai's STT and TTS models?
Pulse measures 5.3% word error rate (15th of 30) among STT systems. Last measured 2026-09-15. Lightning v3.1 Pro measures 4.4% word error rate (5th of 28) among TTS systems. Last measured 2026-09-15.
Which Smallest.ai model is fastest?
Its fastest dated STT result is Pulse at mean 212 ms time to final segment (16th of 28) among STT systems, with 5.3% WER. Last measured 2026-09-15. Its fastest dated TTS result is Lightning v3.1 Pro at mean 362 ms time to first audio (18th of 28) among TTS systems, with 4.4% WER. Last measured 2026-09-15.
Limits of this comparison
Coval measures Pulse on incoming audio and Lightning v3.1 Pro on fixed text prompts.
- Coval does not publish one combined Smallest.ai pipeline score or evaluate every voice-cloning capability.
Official resources
Same datasets, prompts and metric definitions for every model, measured by Coval’s open-source runner. Full methodology on the overview.
Evaluate your own voice agent
Use Coval to test your production configuration, prompts and calls—not only the public benchmark endpoints.