Ga11y: An Automated GIF Annotation System for Visually Impaired Users
Mingrui Ray Zhang, Mingyuan Zhong, Jacob O. Wobbrock
Abstract
Animated GIF images have become prevalent in internet culture, often used to express richer and more nuanced meanings than static images. But animated GIFs often lack adequate alternative text descriptions, and it is challenging to generate such descriptions automatically, resulting in inaccessible GIFs for blind or low-vision (BLV) users. To improve the accessibility of animated GIFs for BLV users, we provide a system called Ga11y (pronounced “galley”), for creating GIF annotations. Ga11y combines the power of machine intelligence and crowdsourcing and has three components: an Android client for submitting annotation requests, a backend server and database, and a web interface where volunteers can respond to annotation requests. We evaluated three human annotation interfaces and employ the one that yielded the best annotation quality. We also conducted a multi-stage evaluation with 12 BLV participants from the United States and China, receiving positive feedback.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b17a755c-b520-4a59-bb29-e959f3d2b9b9Cited by top-tier papers5
- OmniScribe: Authoring Immersive Audio Descriptions for 360° VideosRuei-Che Chang, Chao-Hsien Ting, Chia-Sheng Hung, Wan-Chen Lee et al.UIST 2022 · 32 citations
- DanmuA11y: Making Time-Synced On-Screen Video Comments (Danmu) Accessible to Blind and Low Vision Users via Multi-Viewer Audio DiscussionsShuchang Xu, Xiaofu Jin, Huamin Qu, Yukang YanCHI 2025 · 26 citations
- Front Row: Automatically Generating Immersive Audio Representations of Tennis Broadcasts for Blind ViewersGaurav Jain, Basel Hindi, Connor Courtien, Xin Yi Therese Xu et al.UIST 2023 · 16 citations
- CoSight: Exploring Viewer Contributions to Online Video Accessibility Through Descriptive CommentingRuolin Wang, Xingyu Bruce Liu, Biao Wang, Wayne Zhang et al.UIST 2025 · 1 citation
- "My Brother Is a School Principal, Earns About $80, 000 Per Year... But When the Kids See Me, 'Wow, Uncle, You Have 1, 500 Followers on TikTok!'": A Study of Blind TikTokers' Alternative Professional Development ExperiencesYao Lyu, Tawanna R. Dillahunt, Jiaying Liu, John M. CarrollCHI 2026 · 1 citation
Builds on3
- Revamp: Enhancing Accessible Information Seeking Experience of Online Shopping for Blind or Low Vision UsersRuolin Wang, Zixuan Chen, Mingrui Ray Zhang, Zhaoheng Li et al.CHI 2021 · 37 citations
- "I Hope This Is Helpful": Understanding Crowdworkers' Challenges and Motivations for an Image Description TaskRachel N. Simons, Danna Gurari, Kenneth R. FleischmannCSCW 2020 · 27 citations
- Voicemoji: Emoji Entry Using Voice for Visually Impaired PeopleMingrui Ray Zhang, Ruolin Wang, Xuhai Xu, Qisheng Li et al.CHI 2021 · 25 citations
Related papers
- VideoA11y: Method and Dataset for Accessible Video DescriptionChaoyu Li, Sid Padmanabhuni, Maryam S. Cheema, Hasti Seifi et al.CHI 2025 · 23 citations
- A11yBoard: Making Digital Artboards Accessible to Blind and Low-Vision UsersZhuohao Jerry Zhang, Jacob O. WobbrockCHI 2023 · 25 citations
- Twitter A11y: A Browser Extension to Make Twitter Images AccessibleCole Gleason, Amy Pavel, Emma McCamey, Christina Low et al.CHI 2020 · 123 citations
- Say It All: Feedback for Improving Non-Visual Presentation AccessibilityYi-Hao Peng, JiWoong Jang, Jeffrey P. Bigham, Amy PavelCHI 2021 · 46 citations
- SPICA: Interactive Video Content Exploration through Augmented Audio Descriptions for Blind or Low-Vision ViewersZheng Ning, Brianna L. Wimer, Kaiwen Jiang, Keyi Chen et al.CHI 2024 · 27 citations
