DATASETTTSCOVAL PROMPT MANIFESTACTIVE

manifest

Text prompts text-to-speech benchmark dataset

30 short text prompts that every text-to-speech model synthesizes.

Items
30
fixed public inputs
Models measured
28
last 30 days
Current leader · TTFA#1 / 28
66ms
vuivia Fluxions

How models rank on Text prompts

Full TTS dashboard
Time to First Audio on Text promptsMilliseconds · lower is better · 30-day average on this datasetEvery TTS model measured on the Text prompts dataset, ranked on Time to First Audio.
  1. #1vui66 ms
  2. #3TTS Flash 291 ms
  3. #4Qwen3 TTS 1.7b106 ms
  4. #6TTS 2181 ms
  5. #7Flash v2.5194 ms
  6. #8Default235 ms
  7. #9TTS Rt v2245 ms
  8. #10TTS RT v1245 ms
  9. #11Mist v3258 ms
  10. #12Sonic 3.5277 ms
Show all 28 models
  1. #14Aura 2308 ms
  2. #15Coda312 ms
  3. #16S2.1 Pro335 ms
  4. #19S1379 ms
  5. #20Grok TTS397 ms
  6. #21Sonic 3.6423 ms
  7. #22Simba 3.2451 ms
  8. #23Simba 3.0472 ms
  9. #24Chirp 3 HD535 ms
  10. #25Falcon 2545 ms
  11. #27S2.1 Pro Free963 ms
  12. #28GPT-4o mini TTS1013 ms
median of all models · 310 ms
TTS models on the Text prompts dataset over the last 30 days, ranked on Time to First Audio.
#ModelHostTTFAWERSamples
1vuiFluxions66 ms9,490
2Qwen3 TTS FastNari71 ms2,230
3TTS Flash 2Inworld AI91 ms11,441
4Qwen3 TTS 1.7bBaseten106 ms420
5Palabra TTS v1Palabra115 ms11,425
6TTS 2Inworld AI181 ms11,441
7Flash v2.5ElevenLabs194 ms9,158
8DefaultGradium235 ms9,167
9TTS Rt v2Soniox245 ms11,232
10TTS RT v1Soniox245 ms11,234
11Mist v3Rime258 ms11,427
12Sonic 3.5Cartesia277 ms11,440
13Phantom Z 3.4 conversationalDeepdub286 ms11,436
14Aura 2Deepgram308 ms11,426
15CodaRime312 ms11,432
16S2.1 ProFish Audio335 ms11,440
17Eleven v3 ConversationalElevenLabs348 ms11,429
18Lightning v3.1 ProSmallest362 ms11,441
19S1Fish Audio379 ms11,433
20Grok TTSxAI397 ms11,426
21Sonic 3.6Cartesia423 ms8,250
22Simba 3.2Speechify451 ms11,433
23Simba 3.0Speechify472 ms11,436
24Chirp 3 HDGoogle535 ms11,440
25Falcon 2Murf545 ms11,437
26Qwen3 TTS Flash RealtimeAlibaba753 ms11,266
27S2.1 Pro FreeFish Audio963 ms11,406
28GPT-4o mini TTSOpenAI1013 ms11,410
Under-sampled models are excluded; tied models share a place. Dotted WER values split into substitutions, deletions and insertions on hover or tap.

Inside the dataset

30 fixed public inputs — every model is tested on exactly these.

Every prompt in the Text prompts dataset, verbatim from the manifest.
A1Your order #ORD-24589 shipped via UPS tracking 1Z999AA1234567890 and will arrive between 2:30 PM and 4:45 PM tomorrow.
A2There's a slight delay with your $347.89 order, but we expect it to ship by Friday afternoon.
A3Order status update: your items are now in "Quality Check" and should ship within 24-48 hours.
A4Hi Ms. Garcia, your appointment with Dr. Peterson is scheduled for Tuesday, March 5th at 10:30 AM.
A5I can reschedule you to either 11:15 AM on Wednesday or, let me check, 2:45 PM on Thursday.
A6Please arrive 10 minutes early and bring your insurance card and a photo ID.
A7Let's fix your Wi-Fi: unplug your router for exactly 30 seconds, then wait for the green light.
A8Your device is downloading update version 12.4.1, about 250 MB remaining, roughly 8 minutes left.
A9I'm seeing error code E-1047 here, which usually means, actually, let me guide you through the solution.
A10For security, please confirm the last 4 digits of the card ending in 8429 used for your recent purchase.
A11I need to verify your account: what's the ZIP code we have on file? Please take your time.
A12We detected 2 unusual login attempts from Chicago, Illinois yesterday around 11:30 PM.
A13Hi David, your premium subscription renews in 5 days at $29.99, would you like to continue or explore other options?
A14Based on your purchase history, you might like our flash sale: 30% off electronics through Sunday.
A15I noticed you haven't logged in since January 15th, is there anything we can help you with?
A16Your invoice shows a balance of $187.50, including the monthly fee plus tax.
A17The $24.99 charge is for the premium support package you activated on January 22nd.
A18Your auto-payment of $67.50 failed due to an expired card, would you like to update your payment method?
A19Thanks for your interest in our software, companies your size typically see 20-25% efficiency gains.
A20For a 50-person team, I'd recommend our Business plan at $149 per month rather than Basic.
A21Great timing! We're offering 20% off annual plans, that saves you $359.88 for the first year.
A22On a scale of 1 to 10, how would you rate your current pain level?
A23Have you taken any medications today? Please include prescription drugs and over-the-counter items.
A24Your symptoms sound manageable, but I'd recommend seeing Dr. Martinez within 48 hours to be safe.
A25Your PTO request for December 20th through 24th has been approved, that's 3 business days.
A26Password reset complete: your temporary password was sent to your email. Please change it within 24 hours.
A27IT ticket about your printer issue has been assigned to technician Lisa Chen.
A28Your wire transfer of $1,247.50 to account ending in 5691 was processed at 2:15 PM.
A29Hi Mr. O'Connor, I'm calling about your consultation, are you available for 20 minutes between 9 AM and noon?
A30System alert: your server usage hit 87% at 3:42 AM, but it's back to normal levels now.

Source: Coval TTS test cases · License: Apache-2.0

What it tests

Every TTS model synthesizes the same 30 prompts. Time to First Audio measures startup latency on identical text, while transcription-based WER checks whether the generated speech preserves the prompt's words.

Metrics reported for this dataset

Same datasets, prompts and metric definitions for every model, measured by Coval’s open-source runner. Full methodology on the overview.

Evaluate your own voice agent

Use Coval to test your production configuration, prompts and calls—not only the public benchmark endpoints.

Book a Demo