Gradium AI的新TTS模型在速度和准确度上都有出色表现,首音频生成仅需216毫秒。
Gradium AI推出新默认TTS模型,在五种语言的500个困难句子测试中达到81.0%人类评估通过率。该模型在Coval基准上实现216毫秒P50首音频生成时间。评估数据集已以CC BY 4.0许可在Hugging Face开放。
Gradium AI Releases New Default TTS Model: 81.0% Hard-Case Pass Rate at 216 ms Time-to-First-Audio
Speed and accuracy usually pull against each other in text-to-speech. Gradium AI's new default model reports both: an 81.0% human-rated pass rate on 500 hard sentences across five languages, at 216 ms P50 time-to-first-audio on Coval. The evaluation set is open on Hugging Face under CC BY 4.0. The post Gradium AI Releases New Default TTS Model: 81.0% Hard-Case Pass Rate at 216 ms Time-to-First-Audio appeared first on MarkTechPost .