Published event
Research ModelRelease 1 source(s)

Adding Benchmaxxer Repellant to the Open ASR Leaderboard

Updated September 26, 2026 · 2:44 PM · source date May 6, 2026

Summary

Adding Benchmaxxer Repellant to the Open ASR Leaderboard Adding Benchmaxxer Repellant to the Open ASR Leaderboard Published May 6, 2026 Update on GitHub Upvote 20 Eric Bezzam bezzam Steven Zheng Steveeeeeeen Eustache Le Bihan eustlb Sergio Bruccoleri SBruccoleriAppen AppenAIResearch Jeanine Sinanan-Singh jmss-appen AppenAIResearch Casey Ford c-e-ford-appen AppenAIResearch Guanbo Wang wgb14 DataoceanAI1 Yukai Huang YukaiHuang DataoceanAI1 Ke Li like2026 DataoceanAI1 Yufeng Hao logicbean DataoceanAI1 Liao Xiaoling ally-lxl DataoceanAI1 "When a measure becomes a target, it ceases to be a good measure." (Goodhart’s Law) TLDR : Appen Inc. and DataoceanAI have provided high-quality English ASR datasets covering scripted and conversational speech over multiple accents.

Why it matters

This ModelRelease is relevant to the technology intelligence record because it involves GitHub, Hugging Face. The source article should remain the factual reference for follow-up coverage.

Key facts
  • Adding Benchmaxxer Repellant to the Open ASR Leaderboard Published May 6, 2026 Update on GitHub Upvote 20 Eric Bezzam bezzam Steven Zheng Steveeeeeeen Eustache Le Bihan eustlb Sergio Bruccoleri SBruccoleriAppen AppenAIResearch Jeanine Sinanan-Singh jmss-appen AppenAIResearch Casey Ford c-e-ford-appen AppenAIResearch Guanbo Wang wgb14 DataoceanAI1 Yukai Huang YukaiHuang DataoceanAI1 Ke Li like2026 DataoceanAI1 Yufeng Hao logicbean DataoceanAI1 Liao Xiaoling ally-lxl DataoceanAI1 "When a measure becomes a target, it ceases to be a good measure." (Goodhart’s Law) TLDR : Appen Inc.
  • and DataoceanAI have provided high-quality English ASR datasets covering scripted and conversational speech over multiple accents.
  • To prevent potential risks of benchmaxxing or test-set contamination, we will keep these datasets private for a high-quality measure of performance on multiple tasks.
  • We’re not updating the average WER at this time : by default, the leaderboard’s Average WER remains computed on public datasets only.
  • You can optionally include the private datasets using the toggle to see their impact 👀 Since its launch in September 2023, the Open ASR Leaderboard has been visited over 710K times.
  • We’re blown away by the community’s interest and motivation to keep pushing speech recognition 🗣️ Two words sum up the objectives (but also challenges) in maintaining a benchmark like the Open ASR Leaderboard: Standardization : models can have different conventions for their usage and outputs, e.g.
Entities in this story
Related events