PROVIDERLAST 30 DAYS
MiniMax voice AI models and benchmarks
MiniMax develops generative audio models and hosts the measured Speech 2.8 HD and Turbo synthesis endpoints.
- Measured models
- 2
- TTS
Overview
Within the Speech 2.8 family, HD targets quality while Turbo targets speed.
Every model below is measured daily on the same fixed inputs and ranked against the full field, never blended into a company score.
Model lineup
- Text-to-Speech
- Speech 2.8 HD, Speech 2.8 Turbo
Text-to-Speech
Full TTS dashboardRanked on Time to First Audio against 30 measured models.
leaders plus MiniMax models · 21 other models in the full table
- #6Default4.6%
leaders plus MiniMax models · 21 other models in the full table
| Model | Host | TTFA | Rank |
|---|---|---|---|
| Speech 2.8 HD | MiniMax | 460 ms | 21st of 30 |
| Speech 2.8 Turbo | MiniMax | 411 ms | 16th of 30 |
Limits of this comparison
- Coval does not infer subjective quality from `HD` or speed from `Turbo`, and it does not score cloning or emotion fidelity.
Official resources
- MiniMax Speech API (documentation)
Same datasets, prompts and metric definitions for every model, measured by Coval’s open-source runner. Full methodology on the overview.
Evaluate your own voice agent
Use Coval to test your production configuration, prompts and calls—not only the public benchmark endpoints.