Lost in Transcription: Subtitle Errors in Automatic Speech Recognition Reduce Speaker and Content Evaluations
Kowe Kadoma, Priyal Shrivastava, Mor Naaman
Abstract
Researchers have demonstrated that Automatic Speech Recognition (ASR) systems perform differently across demographic groups. In this work, we examined how subtitle errors affect evaluations of speakers and their content using a preregistered online experiment (N=207, U.S.-based crowdworkers). Participants watched speakers with various accents deliver a talk in which the subtitles were accurate or error-prone. Our results indicate that error-prone subtitles consistently reduce both speaker and content evaluations for all speakers. We did not see disparate impact between the accent groups, controlling for subtitle quality. Taken together, though, the findings of this short paper imply that speakers with accents for which ASR systems perform poorly are likely to be further penalized by viewers with lower evaluations.
CCS Concepts: • Human-centered computing → Empirical studies in HCI; HCI design and evaluation methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext bedebb9b-d9c2-4238-bfe5-c1e0fe5c7031Builds on8
- "It's Kind of Like Code-Switching": Black Older Adults' Experiences with a Voice Assistant for Health Information SeekingChristina N. Harrington, Radhika Garg, Amanda T. Woodward, Dimitri WilliamsCHI 2022 · 99 citations
- Can Voice Assistants Be Microaggressors? Cross-Race Psychological Responses to Failures of Automatic Speech RecognitionKimi Wenzel, Nitya Devireddy, Cam Davidson, Geoff KaufmanCHI 2023 · 24 citations
- Designing for Harm Reduction: Communication Repair for Multicultural Users' Voice InteractionsKimi Wenzel, Geoff KaufmanCHI 2024 · 17 citations
- How Users Experience Closed Captions on Live Television: Quality Metrics Remain a ChallengeMariana Arroyo Chavez, Molly Feanny, Matthew Seita, Bernard Thompson et al.CHI 2024 · 15 citations
- Toward Language Justice: Exploring Multilingual Captioning for AccessibilityAashaka Desai, Rahaf Alharbi, Stacy Hsueh, Richard E. Ladner et al.CHI 2025 · 11 citations
Related papers
- Is the Same Performance Really the Same?: Understanding How Listeners Perceive ASR Results Differently According to the Speaker's AccentSeoyoung Kim, Yeon Su Park, Dakyeom Ahn, Jin Myung Kwak et al.CSCW 2024 · 3 citations
- "It feels like we're not meeting the criteria": Examining and Mitigating the Cascading Effects of Bias in Automatic Speech Recognition in Spoken Language InterfacesKelechi Ezema, Chelsea Chandler, Rosy Southwell, Niranjan Cholendiran et al.CHI 2025 · 8 citations
- Impact of Annotator Demographics on Sentiment Dataset LabelingYi Ding, Jacob You, Tonja-Katrin Machulla, Jennifer Jacobs et al.CSCW 2022 · 20 citations
- Adaptive Subtitles: Preferences and Trade-Offs in Real-Time Media AdaptionBenjamin M. Gorman, Michael Crabb, Michael ArmstrongCHI 2021 · 31 citations
- A View on the Viewer: Gaze-Adaptive Captions for VideosKuno Kurzhals, Fabian Göbel, Katrin Angerbauer, Michael Sedlmair et al.CHI 2020 · 42 citations
