Text prompts text-to-speech benchmark dataset
30 short text prompts that every text-to-speech model synthesizes.
- Items
- 30
- fixed public inputs
- Models measured
- 30
- last 30 days
How models rank on Text prompts
Full TTS dashboardShow all 30 modelsShow fewer
- #15Default384 ms
median of all models · 397 ms
- #6Default4.6%
Show all 30 modelsShow fewer
median of all models · 4.9%
| # | Model | Host | TTFA | WER | Samples |
|---|---|---|---|---|---|
| 1 | Palabra TTS v1 | Palabra | 116 ms | 5.9% | 14,396 |
| 2 | vui | Fluxions | 124 ms | 8,624 | |
| 3 | TTS Flash 2 | Inworld AI | 128 ms | 8,719 | |
| 4 | TTS 2 | Inworld AI | 176 ms | 4.8% | 14,529 |
| 5 | Blizzard | Lmnt | 235 ms | 7.4% | 14,117 |
| 6 | Neural | Azure | 236 ms | 4.3% | 3,350 |
| 7 | Mist v3 | Rime | 256 ms | 6.3% | 14,524 |
| 8 | TTS Rt v2 | Soniox | 262 ms | 8,849 | |
| 9 | Sonic 3.5 | Cartesia | 274 ms | 6.1% | 14,513 |
| 10 | TTS RT v1 | Soniox | 275 ms | 3.7% | 14,528 |
| 11 | Dragon HD Latest | Azure | 310 ms | 5.1% | 3,350 |
| 12 | Coda | Rime | 313 ms | 5.2% | 14,523 |
| 13 | Aura 2 | Deepgram | 328 ms | 5.3% | 14,501 |
| 14 | S2.1 Pro | Fish Audio | 374 ms | 10,222 | |
| 15 | Default | Gradium | 384 ms | 4.6% | 14,002 |
| 16 | Speech 2.8 Turbo | MiniMax | 411 ms | 4.8% | 1,997 |
| 17 | Eleven v3 Conversational | ElevenLabs | 412 ms | 6,750 | |
| 18 | Grok TTS | xAI | 420 ms | 4.7% | 14,524 |
| 19 | S1 | Fish Audio | 434 ms | 4.9% | 10,236 |
| 20 | Flash v2.5 | ElevenLabs | 455 ms | 6.7% | 9,259 |
| 21 | Speech 2.8 HD | MiniMax | 460 ms | 1,995 | |
| 22 | Sonic 3.6 | Cartesia | 466 ms | 560 | |
| 23 | Simba 3.2 | Speechify | 484 ms | 4.7% | 14,525 |
| 24 | Chirp 3 HD | 512 ms | 5.2% | 14,530 | |
| 25 | Simba 3.0 | Speechify | 526 ms | 5.3% | 14,526 |
| 26 | Falcon 2 | Murf | 549 ms | 10,229 | |
| 27 | Lightning v3.1 Pro | Smallest | 586 ms | 4.4% | 14,515 |
| 28 | Qwen3 TTS Flash Realtime | Alibaba | 645 ms | 8.8% | 14,530 |
| 29 | S2.1 Pro Free | Fish Audio | 806 ms | 4.7% | 14,464 |
| 30 | GPT-4o mini TTS | OpenAI | 1075 ms | 4.8% | 14,517 |
Inside the dataset
30 fixed public inputs — every model is tested on exactly these.
| A1 | Your order #ORD-24589 shipped via UPS tracking 1Z999AA1234567890 and will arrive between 2:30 PM and 4:45 PM tomorrow. |
| A2 | There's a slight delay with your $347.89 order, but we expect it to ship by Friday afternoon. |
| A3 | Order status update: your items are now in "Quality Check" and should ship within 24-48 hours. |
| A4 | Hi Ms. Garcia, your appointment with Dr. Peterson is scheduled for Tuesday, March 5th at 10:30 AM. |
| A5 | I can reschedule you to either 11:15 AM on Wednesday or, let me check, 2:45 PM on Thursday. |
| A6 | Please arrive 10 minutes early and bring your insurance card and a photo ID. |
| A7 | Let's fix your Wi-Fi: unplug your router for exactly 30 seconds, then wait for the green light. |
| A8 | Your device is downloading update version 12.4.1, about 250 MB remaining, roughly 8 minutes left. |
| A9 | I'm seeing error code E-1047 here, which usually means, actually, let me guide you through the solution. |
| A10 | For security, please confirm the last 4 digits of the card ending in 8429 used for your recent purchase. |
| A11 | I need to verify your account: what's the ZIP code we have on file? Please take your time. |
| A12 | We detected 2 unusual login attempts from Chicago, Illinois yesterday around 11:30 PM. |
| A13 | Hi David, your premium subscription renews in 5 days at $29.99, would you like to continue or explore other options? |
| A14 | Based on your purchase history, you might like our flash sale: 30% off electronics through Sunday. |
| A15 | I noticed you haven't logged in since January 15th, is there anything we can help you with? |
| A16 | Your invoice shows a balance of $187.50, including the monthly fee plus tax. |
| A17 | The $24.99 charge is for the premium support package you activated on January 22nd. |
| A18 | Your auto-payment of $67.50 failed due to an expired card, would you like to update your payment method? |
| A19 | Thanks for your interest in our software, companies your size typically see 20-25% efficiency gains. |
| A20 | For a 50-person team, I'd recommend our Business plan at $149 per month rather than Basic. |
| A21 | Great timing! We're offering 20% off annual plans, that saves you $359.88 for the first year. |
| A22 | On a scale of 1 to 10, how would you rate your current pain level? |
| A23 | Have you taken any medications today? Please include prescription drugs and over-the-counter items. |
| A24 | Your symptoms sound manageable, but I'd recommend seeing Dr. Martinez within 48 hours to be safe. |
| A25 | Your PTO request for December 20th through 24th has been approved, that's 3 business days. |
| A26 | Password reset complete: your temporary password was sent to your email. Please change it within 24 hours. |
| A27 | IT ticket about your printer issue has been assigned to technician Lisa Chen. |
| A28 | Your wire transfer of $1,247.50 to account ending in 5691 was processed at 2:15 PM. |
| A29 | Hi Mr. O'Connor, I'm calling about your consultation, are you available for 20 minutes between 9 AM and noon? |
| A30 | System alert: your server usage hit 87% at 3:42 AM, but it's back to normal levels now. |
Source: Coval TTS test cases · License: Apache-2.0
What it tests
Every TTS model synthesizes the same 30 prompts. Time to First Audio measures startup latency on identical text, while transcription-based WER checks whether the generated speech preserves the prompt's words.
Metrics reported for this dataset
Same datasets, prompts and metric definitions for every model, measured by Coval’s open-source runner. Full methodology on the overview.
Evaluate your own voice agent
Use Coval to test your production configuration, prompts and calls—not only the public benchmark endpoints.