Speech technology penalizes some voices: recognition errs nearly twice as often for Black speakers, and accuracy declines for second-language accents and older speakers. We introduce TRIAD, an audit g...
arXiv:2609.18533v1 Announce Type: new
Abstract: Automatic speech recognition (ASR) systems exhibit unequal error rates across speaker groups, motivating interventions on their internal representation...
By Nicolas Bourrel, Abderrahmane Issam, Gerasimos Spanakis
Automatic speech recognition (ASR) systems exhibit unequal error rates across speaker groups, motivating interventions on their internal representations. We ask whether speaker-linked attributes that...
arXiv:2609.38106v1 Announce Type: cross
Abstract: Speech-LLMs are expensive to run, making compression important for real-world deployment. However, compressed models are usually selected using aggre...
By Ganesh Pavan Kartikeya Bharadwaj Kolluri, Michael Kampouridis, Ravi Shekhar
arXiv:2608.30853v1 Announce Type: cross
Abstract: While automatic speech recognition (ASR) models have achieved remarkable improvements in recent years, performance disparities persist across differe...
By Ting-Hui Cheng, Line Katrine Harder Clemmensen, Sneha Das
arXiv:2609.36500v1 Announce Type: cross
Abstract: Speaker verification systems encounter combinations of noise, channel distortion, and changes in speech. Evaluating each condition separately does no...
By Kamel Kamel, Hridoy Sankar Dutta, Keshav Sood, Sunil Aryal