Watch It, Don't Imagine It: Creating a Better Caption-Occlusion Metric by Collecting More Ecologically Valid Judgments from DHH Viewers
Akhter Al Amin, Saad Hassan, Sooyeon Lee, Matt Huenerfauth
Abstract
Television captions blocking visual information causes dissatisfaction among Deaf and Hard of Hearing (DHH) viewers, yet existing caption evaluation metrics do not consider occlusion. To create such a metric, DHH participants in a recent study imagined how bad it would be if captions blocked various on-screen text or visual content. To gather more ecologically valid data for creating an improved metric, we asked 24 DHH participants to give subjective judgments of caption quality after actually watching videos, and a regression analysis revealed which on-screen contents’ occlusion related to users’ judgments. For several video genres, a metric based on our new dataset out-performed the prior state-of-the-art metric for predicting the severity of captions occluding content during videos, which had been based on that prior study. We contribute empirical findings for improving DHH viewers’ experience, guiding the placement of captions to minimize occlusions, and automated evaluation of captioning quality in television broadcasts.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Cited by top-tier papers5
- "Caption It in an Accessible Way That Is Also Enjoyable": Characterizing User-Driven Captioning Practices on TikTokEmma J. McDonnell, Tessa Eagle, Pitch Sinlapanuntakul, Soo Hyun Moon et al.CHI 2024 · 33 citations
- "Easier or Harder, Depending on Who the Hearing Person Is": Codesigning Videoconferencing Tools for Small Groups with Mixed Hearing StatusEmma J. McDonnell, Soo Hyun Moon, Lucy Jiang, Steven M. Goodman et al.CHI 2023 · 26 citations
- Toward Language Justice: Exploring Multilingual Captioning for AccessibilityAashaka Desai, Rahaf Alharbi, Stacy Hsueh, Richard E. Ladner et al.CHI 2025 · 11 citations
- Unspoken Sound: Identifying Trends in Non-Speech Audio Captioning on YouTubeLloyd May, Keita Ohshiro, Khang Dang, Sripathi Sridhar et al.CHI 2024 · 10 citations
- PeriphAR: Fast and Accurate Real-World Object Selection with Peripheral Augmented Reality DisplaysYutong Ren, Arnav Reddy, Michael NebelingCHI 2026 · 2 citations
Related papers
- How Users Experience Closed Captions on Live Television: Quality Metrics Remain a ChallengeMariana Arroyo Chavez, Molly Feanny, Matthew Seita, Bernard Thompson et al.CHI 2024 · 15 citations
- OnomaCap: Making Non-speech Sound Captions Accessible and Enjoyable through Onomatopoeic Sound RepresentationJooYeong Kim, Jin-Hyuk HongCHI 2025 · 7 citations
- Visualization of Speech Prosody and Emotion in Captions: Accessibility for Deaf and Hard-of-Hearing UsersCaluã de Lacerda Pataca, Matthew Watkins, Roshan L. Peiris, Sooyeon Lee et al.CHI 2023 · 39 citations
- Redesigning Educational Videos for Deaf and Hard-of-Hearing LearnersSi Chen, Haocong Cheng, Suzy Su, Lu Ming et al.CHI 2026 · 1 citation
- Visible Nuances: A Caption System to Visualize Paralinguistic Speech Cues for Deaf and Hard-of-Hearing IndividualsJooYeong Kim, Sooyeon Ahn, Jin-Hyuk HongCHI 2023 · 24 citations
