跳到正文
原文
Hugging Face Blog·· 2026-08-21AI 评分44

Hugging Face 研究:11 款开源 ASR 模型存在基准测试“刷分”现象

Measuring benchmark optimization in speech recognition

AI 导读

Hugging Face 最新研究用三项测试量化语音识别中的“benchmaxxing”,评估 11 款开源 ASR 模型后发现,多款高分模型会复现 VoxPopuli 英文集和 LibriSpeech(clean、other)的基准转录文本,即使音频与之矛盾或相关词已被静音。

来源:Hugging Face Blog · huggingface.co