Responsible Benchmarking of Fairness for Automatic Speech Recognition

Felix E. Herron, Ange Richard, François Portet, Alexandre Allauzen, Solange Rossato


Abstract
Many studies have shown automatic speech processing (ASR) systems have unequal performance across speaker groups (SG’s). However, the manner in which such studies arrive at this conclusion is inconsistent. To pave the way for more reliable results in future studies, we lay out best practices for benchmarking ASR fairness based on literature from machine learning fairness, social sciences, and speech science. We then perform a case study on the Fair-speech benchmark, applying aforementioned best practices, and discuss how failing to do so can result in erroneous conclusions. On the whole, we advocate for as fine-grained an analysis as possible, taking into account as many variables as are available, in order to eschew dataset-level bias.
Anthology ID:
2026.speakable-1.8
Volume:
Proceedings of Speech Language Models in Low-Resource Settings: Performance, Evaluation, and Bias Analysis (SPEAKABLE) @ LREC 2026
Month:
May
Year:
2026
Address:
Palma, Mallorca (Spain)
Editors:
Nina Hosseini-Kivanani, Alessio Brutti, Marco Matassoni, Sandipana Dowerah, Davide Liga, Christoph Schommer
Venues:
SPEAKABLE | WS
SIG:
Publisher:
ELRA Language Resources Association (ELRA)
Note:
Pages:
66–78
Language:
External URL:
https://lrec.elra.info/lrec2026-ws-speakable-08
DOI:
10.63317/3jpc2uj4pp3g
Bibkey:
Cite (ACL):
Felix E. Herron, Ange Richard, François Portet, Alexandre Allauzen, and Solange Rossato. 2026. Responsible Benchmarking of Fairness for Automatic Speech Recognition. In Proceedings of Speech Language Models in Low-Resource Settings: Performance, Evaluation, and Bias Analysis (SPEAKABLE) @ LREC 2026, pages 66–78, Palma, Mallorca (Spain). ELRA Language Resources Association (ELRA).
Cite (Informal):
Responsible Benchmarking of Fairness for Automatic Speech Recognition (Herron et al., SPEAKABLE 2026)
Copy Citation: