"It feels like we're not meeting the criteria": Examining and Mitigating the Cascading Effects of Bias in Automatic Speech Recognition in Spoken Language Interfaces
Kelechi Ezema, Chelsea Chandler, Rosy Southwell, Niranjan Cholendiran, Sidney D'Mello
摘要
Researchers have demonstrated that Automatic Speech Recognition (ASR) systems perform differently across demographic groups (i.e. show bias), yet their downstream impact on spoken language interfaces remains unexplored. We examined this question in the context of a real-world AI-powered interface that provides tutors with feedback on the quality of their discourse. We found that the Whisper ASR had lower accuracy for Black vs. white tutors, likely due to differences in acoustic patterns of speech. The downstream automated discourse classifiers of tutor talk were correspondingly less accurate for Black tutors when presented with ASR input. As a result, although Black tutors demonstrated higher-quality discourse on human transcripts, this trend was not evident on ASR transcripts. We experimented with methods to reduce ASR bias, finding that fine-tuning the ASR on Black speech reduced, but did not eliminate, ASR bias and its downstream effects. We discuss implications for AI-based spoken language interfaces aimed at providing unbiased assessments to improve performance outcomes.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Challenges in Automatic Speech Recognition for Adults with Cognitive ImpairmentMichelle Cohn, Alyssa Lanzi, Yui Ishihara, Chen-Nee Chuah 等CHI 2026 · 被引用 2 次
- Gamifying Compassion: Mitigating Dialect Prejudice Through An AI-Driven Serious GameSicheng Lu, Erick Purwanto, Hong Liu, Adel Chaouch-Orozco 等CHI 2026 · 被引用 1 次
- Lost in Transcription: Subtitle Errors in Automatic Speech Recognition Reduce Speaker and Content EvaluationsKowe Kadoma, Priyal Shrivastava, Mor NaamanCHI 2026 · 被引用 1 次
- Listening Like Humans: Semantics-Guided Noise-Robust Multimodal Speech RecognitionYan Fang, Jun Chen, Yian Yao, Shuxin Zhong 等ACL 2026
它引用的顶会 Paper8
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu 等ICLR 2022 · 被引用 18,833 次
- Robust Speech Recognition via Large-Scale Weak SupervisionAlec Radford, Jong Wook Kim, Tao Xu, Greg Brockman 等ICML 2023 · 被引用 6,966 次
- Critical Race Theory for HCIIhudiya Finda Ogbonnaya-Ogburu, Angela D. R. Smith, Alexandra To, Kentaro ToyamaCHI 2020 · 被引用 397 次
- "It's Kind of Like Code-Switching": Black Older Adults' Experiences with a Voice Assistant for Health Information SeekingChristina N. Harrington, Radhika Garg, Amanda T. Woodward, Dimitri WilliamsCHI 2022 · 被引用 99 次
- Toward Automated Feedback on Teacher Discourse to Enhance Teacher LearningEmily Jensen, Meghan Dale, Patrick J. Donnelly, Cathlyn Stone 等CHI 2020 · 被引用 90 次
相关 Paper
- Can Voice Assistants Be Microaggressors? Cross-Race Psychological Responses to Failures of Automatic Speech RecognitionKimi Wenzel, Nitya Devireddy, Cam Davidson, Geoff KaufmanCHI 2023 · 被引用 24 次
- Is the Same Performance Really the Same?: Understanding How Listeners Perceive ASR Results Differently According to the Speaker's AccentSeoyoung Kim, Yeon Su Park, Dakyeom Ahn, Jin Myung Kwak 等CSCW 2024 · 被引用 3 次
- Beyond WER: Probing Whisper's Sub-token Decoder Across Diverse Language Resource LevelsSiyu Liang, Nicolas Ballier, Gina-Anne Levow, Richard A. WrightEMNLP 2025
- Error-preserving Automatic Speech Recognition of Young English Learners' LanguageJanick Michot, Manuela Hürlimann, Jan Deriu, Luzia Sauer 等ACL 2024 · 被引用 2 次
- Self-Taught Recognizer: Toward Unsupervised Adaptation for Speech Foundation ModelsYuchen Hu, Chen Chen, Chao-Han Huck Yang, Chengwei Qin 等NeurIPS 2024 · 被引用 14 次
