StepFun ships five voice models, word error rate hits 1.7%
StepFun launched five StepAudio 3 models in one day, topping Artificial Analysis leaderboards for real-time dialogue and speech reasoning, with ASR hitting a 1.7% word error rate.
4 verified stories covering Speech Recognition, product updates and industry developments.
StepFun launched five StepAudio 3 models in one day, topping Artificial Analysis leaderboards for real-time dialogue and speech reasoning, with ASR hitting a 1.7% word error rate.
Xiaomi's open-source CocktailASR-1 model targets one speaker's voice in overlapping speech, scoring a 12.29% error rate with three people talking at once.
A Hume AI probe of 11 leading speech recognition models finds the top scorers on public benchmarks are also the most likely to echo wrong reference transcripts instead of what they actually hear.
Tencent released Hy ASR 3.0 preview with public WER figures, dialect coverage and Tencent Cloud API limits.