Responsible Benchmarking of Fairness for Automatic Speech Recognition
Felix E. Herron, Ange Richard, François Portet, Alexandre Allauzen, Solange Rossato
Abstract
Many studies have shown automatic speech processing (ASR) systems have unequal performance across speaker groups (SG’s). However, the manner in which such studies arrive at this conclusion is inconsistent. To pave the way for more reliable results in future studies, we lay out best practices for benchmarking ASR fairness based on literature from machine learning fairness, social sciences, and speech science. We then perform a case study on the Fair-speech benchmark, applying aforementioned best practices, and discuss how failing to do so can result in erroneous conclusions. On the whole, we advocate for as fine-grained an analysis as possible, taking into account as many variables as are available, in order to eschew dataset-level bias.- Anthology ID:
- 2026.speakable-1.8
- Volume:
- Proceedings of Speech Language Models in Low-Resource Settings: Performance, Evaluation, and Bias Analysis (SPEAKABLE) @ LREC 2026
- Month:
- May
- Year:
- 2026
- Address:
- Palma, Mallorca (Spain)
- Editors:
- Nina Hosseini-Kivanani, Alessio Brutti, Marco Matassoni, Sandipana Dowerah, Davide Liga, Christoph Schommer
- Venues:
- SPEAKABLE | WS
- SIG:
- Publisher:
- ELRA Language Resources Association (ELRA)
- Note:
- Pages:
- 66–78
- Language:
- External URL:
- https://lrec.elra.info/lrec2026-ws-speakable-08
- DOI:
- 10.63317/3jpc2uj4pp3g
- Cite (ACL):
- Felix E. Herron, Ange Richard, François Portet, Alexandre Allauzen, and Solange Rossato. 2026. Responsible Benchmarking of Fairness for Automatic Speech Recognition. In Proceedings of Speech Language Models in Low-Resource Settings: Performance, Evaluation, and Bias Analysis (SPEAKABLE) @ LREC 2026, pages 66–78, Palma, Mallorca (Spain). ELRA Language Resources Association (ELRA).
- Cite (Informal):
- Responsible Benchmarking of Fairness for Automatic Speech Recognition (Herron et al., SPEAKABLE 2026)