Gradium AI Releases New Default TTS Model: 81.0% Hard-Case Pass Rate at 216 ms Time-to-First-Audio
What changed
Gradium AI released a new default text-to-speech model, which turns written text into spoken audio. The company reports that it passed 81.0% of 500 difficult, human-rated sentences in five languages, with the first audio arriving in a median 216 milliseconds on Coval.
What this means for you
Voice-app developers and creators may be able to deliver faster, more natural-sounding speech with this release. The 500-sentence evaluation set is available on Hugging Face under a Creative Commons license, making the claims easier to review and reuse.
