Hugging Face Blog·· 2026-08-21AI 评分44
Hugging Face 研究:11 款开源 ASR 模型存在基准测试“刷分”现象
Measuring benchmark optimization in speech recognition
AI 导读
Hugging Face 最新研究用三项测试量化语音识别中的“benchmaxxing”,评估 11 款开源 ASR 模型后发现,多款高分模型会复现 VoxPopuli 英文集和 LibriSpeech(clean、other)的基准转录文本,即使音频与之矛盾或相关词已被静音。
来源:Hugging Face Blog · huggingface.co