Lost in Transcription: Subtitle Errors in Automatic Speech Recognition Reduce Speaker and Content Evaluations
Kowe Kadoma, Priyal Shrivastava, Mor Naaman
摘要
Researchers have demonstrated that Automatic Speech Recognition (ASR) systems perform differently across demographic groups. In this work, we examined how subtitle errors affect evaluations of speakers and their content using a preregistered online experiment (N=207, U.S.-based crowdworkers). Participants watched speakers with various accents deliver a talk in which the subtitles were accurate or error-prone. Our results indicate that error-prone subtitles consistently reduce both speaker and content evaluations for all speakers. We did not see disparate impact between the accent groups, controlling for subtitle quality. Taken together, though, the findings of this short paper imply that speakers with accents for which ASR systems perform poorly are likely to be further penalized by viewers with lower evaluations.
CCS Concepts: • Human-centered computing → Empirical studies in HCI; HCI design and evaluation methods.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper8
- "It's Kind of Like Code-Switching": Black Older Adults' Experiences with a Voice Assistant for Health Information SeekingChristina N. Harrington, Radhika Garg, Amanda T. Woodward, Dimitri WilliamsCHI 2022 · 被引用 99 次
- Can Voice Assistants Be Microaggressors? Cross-Race Psychological Responses to Failures of Automatic Speech RecognitionKimi Wenzel, Nitya Devireddy, Cam Davidson, Geoff KaufmanCHI 2023 · 被引用 24 次
- Designing for Harm Reduction: Communication Repair for Multicultural Users' Voice InteractionsKimi Wenzel, Geoff KaufmanCHI 2024 · 被引用 17 次
- How Users Experience Closed Captions on Live Television: Quality Metrics Remain a ChallengeMariana Arroyo Chavez, Molly Feanny, Matthew Seita, Bernard Thompson 等CHI 2024 · 被引用 15 次
- Toward Language Justice: Exploring Multilingual Captioning for AccessibilityAashaka Desai, Rahaf Alharbi, Stacy Hsueh, Richard E. Ladner 等CHI 2025 · 被引用 11 次
相关 Paper
- Is the Same Performance Really the Same?: Understanding How Listeners Perceive ASR Results Differently According to the Speaker's AccentSeoyoung Kim, Yeon Su Park, Dakyeom Ahn, Jin Myung Kwak 等CSCW 2024 · 被引用 3 次
- "It feels like we're not meeting the criteria": Examining and Mitigating the Cascading Effects of Bias in Automatic Speech Recognition in Spoken Language InterfacesKelechi Ezema, Chelsea Chandler, Rosy Southwell, Niranjan Cholendiran 等CHI 2025 · 被引用 8 次
- Impact of Annotator Demographics on Sentiment Dataset LabelingYi Ding, Jacob You, Tonja-Katrin Machulla, Jennifer Jacobs 等CSCW 2022 · 被引用 20 次
- Adaptive Subtitles: Preferences and Trade-Offs in Real-Time Media AdaptionBenjamin M. Gorman, Michael Crabb, Michael ArmstrongCHI 2021 · 被引用 31 次
- A View on the Viewer: Gaze-Adaptive Captions for VideosKuno Kurzhals, Fabian Göbel, Katrin Angerbauer, Michael Sedlmair 等CHI 2020 · 被引用 42 次
