PROVIDERLAST 30 DAYS

official resource

Rime voice AI models and benchmarks

Rime builds real-time text-to-speech APIs and voices.

Measured models
2
TTS

Overview

Rime supports multilingual and on-premises synthesis across two live generations, Coda and Mist v3.

Every model below is measured daily on the same fixed inputs and ranked against the full field, never blended into a company score.

Model lineup

Text-to-Speech
Coda, Mist v3

Ranked on Time to First Audio against 30 measured models.

Time to First Audioms · lower is better · Rime models markedEvery measured TTS model on Time to First Audio, with Rime's models highlighted.
  1. #2vui124 ms
  2. #3TTS Flash 2128 ms
  3. #4TTS 2176 ms
  4. #5Blizzard235 ms
  5. #6Neural236 ms
  6. #7Mist v3256 ms
  7. #12Coda313 ms
leaders plus Rime models · 22 other models in the full table
Word Error Rate% · lower is better · Rime models markedEvery measured TTS model on Word Error Rate, with Rime's models highlighted.
  1. #1TTS RT v13.7%
  2. #2TTS Rt v23.9%
  3. #3Neural4.3%
  4. #6Default4.6%
  5. #7S2.1 Pro4.6%
  6. #19Coda5.2%
  7. #26Mist v36.3%
leaders plus Rime models · 21 other models in the full table
Benchmarked models with their Time to First Audio over the last 30 days.
ModelHostTTFARank
CodaRime313 ms12th of 30
Mist v3Rime256 ms7th of 30

Limits of this comparison

Coval measures Coda and Mist v3 as separate first-party endpoints using the same fixed prompts.

  • Coval does not rank subjective naturalness or every speaker in Rime's voice catalogue.

Official resources

Same datasets, prompts and metric definitions for every model, measured by Coval’s open-source runner. Full methodology on the overview.

Evaluate your own voice agent

Use Coval to test your production configuration, prompts and calls—not only the public benchmark endpoints.

Book a Demo