Slide Gestalt: Automatic Structure Extraction in Slide Decks for Non-Visual Access
Yi-Hao Peng, Peggy Chi, Anjuli Kannan, Meredith Ringel Morris, Irfan Essa
Abstract
Presentation slides commonly use visual patterns for structural navigation, such as titles, dividers, and build slides. However, screen readers do not capture such intention, making it time-consuming and less accessible for blind and visually impaired (BVI) users to linearly consume slides with repeated content. We present Slide Gestalt, an automatic approach that identifies the hierarchical structure in a slide deck. Slide Gestalt computes the visual and textual correspondences between slides to generate hierarchical groupings. Readers can navigate the slide deck from the higher-level section overview to the lower-level description of a slide group or individual elements interactively with our UI. We derived side consumption and authoring practices from interviews with BVI readers and sighted creators and an analysis of 100 decks. We performed our pipeline with 50 real-world slide decks and a large dataset. Feedback from eight BVI participants showed that Slide Gestalt helped navigate a slide deck by anchoring content more efficiently, compared to using accessible slides.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext bf14ae0e-57ce-4542-abb6-53fde4531cb2Cited by top-tier papers8
- GenAssist: Making Image Generation AccessibleMina Huh, Yi-Hao Peng, Amy PavelUIST 2023 · 58 citations
- CodeA11y: Making AI Coding Assistants Useful for Accessible Web DevelopmentPeya Mowar, Yi-Hao Peng, Jason Wu, Aaron Steinfeld et al.CHI 2025 · 22 citations
- Memory Reviver: Supporting Photo-Collection Reminiscence for People with Visual Impairment via a Proactive ChatbotShuchang Xu, Chang Chen, Zichen Liu, Xiaofu Jin et al.UIST 2024 · 21 citations
- OutlineSpark: Igniting AI-powered Presentation Slides Creation from Computational Notebooks through OutlinesFengjie Wang, Yanna Lin, Leni Yang, Haotian Li et al.CHI 2024 · 19 citations
- The Sky is the Limit: Understanding How Generative AI can Enhance Screen Reader Users' Experience with Productivity ApplicationsMinoli Perera, Swamy Ananthanarayan, Cagatay Goncu, Kim MarriottCHI 2025 · 12 citations
Builds on13
- DiT: Self-supervised Pre-training for Document Image TransformerJunlong Li, Yiheng Xu, Tengchao Lv, Lei Cui et al.ACM MM 2022 · 184 citations
- Twitter A11y: A Browser Extension to Make Twitter Images AccessibleCole Gleason, Amy Pavel, Emma McCamey, Christina Low et al.CHI 2020 · 123 citations
- Rescribe: Authoring and Automatically Editing Audio DescriptionsAmy Pavel, Gabriel Reyes, Jeffrey P. BighamUIST 2020 · 72 citations
- Automatic Generation of Two-Level Hierarchical Tutorials from Instructional Makeup VideosAnh Truong, Peggy Chi, David Salesin, Irfan Essa et al.CHI 2021 · 57 citations
- ImageExplorer: Multi-Layered Touch Exploration to Encourage Skepticism Towards Imperfect AI-Generated Image CaptionsJaewook Lee, Jaylin Herskovitz, Yi-Hao Peng, Anhong GuoCHI 2022 · 55 citations
Related papers
- Understanding Blind Screen-Reader Users' Experiences of Digital ArtboardsAnastasia Schaadhardt, Alexis Hiniker, Jacob O. WobbrockCHI 2021 · 44 citations
- Psychologically-inspired, unsupervised inference of perceptual groups of GUI widgets from GUI imagesMulong Xie, Zhenchang Xing, Sidong Feng, Xiwei Xu et al.FSE 2022 · 30 citations
- Diffscriber: Describing Visual Design Changes to Support Mixed-Ability Collaborative Presentation AuthoringYi-Hao Peng, Jason Wu, Jeffrey P. Bigham, Amy PavelUIST 2022 · 26 citations
- Say It All: Feedback for Improving Non-Visual Presentation AccessibilityYi-Hao Peng, JiWoong Jang, Jeffrey P. Bigham, Amy PavelCHI 2021 · 46 citations
- Enhancing Revisitation in Touchscreen Reading for Visually Impaired People with Semantic Navigation DesignZhichun Li, Yu Jiang, Xiaochen Liu, Yuhang Zhao et al.UbiComp 2022 · 7 citations
