Slide Gestalt: Automatic Structure Extraction in Slide Decks for Non-Visual Access
Yi-Hao Peng, Peggy Chi, Anjuli Kannan, Meredith Ringel Morris, Irfan Essa
摘要
Presentation slides commonly use visual patterns for structural navigation, such as titles, dividers, and build slides. However, screen readers do not capture such intention, making it time-consuming and less accessible for blind and visually impaired (BVI) users to linearly consume slides with repeated content. We present Slide Gestalt, an automatic approach that identifies the hierarchical structure in a slide deck. Slide Gestalt computes the visual and textual correspondences between slides to generate hierarchical groupings. Readers can navigate the slide deck from the higher-level section overview to the lower-level description of a slide group or individual elements interactively with our UI. We derived side consumption and authoring practices from interviews with BVI readers and sighted creators and an analysis of 100 decks. We performed our pipeline with 50 real-world slide decks and a large dataset. Feedback from eight BVI participants showed that Slide Gestalt helped navigate a slide deck by anchoring content more efficiently, compared to using accessible slides.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- GenAssist: Making Image Generation AccessibleMina Huh, Yi-Hao Peng, Amy PavelUIST 2023 · 被引用 58 次
- CodeA11y: Making AI Coding Assistants Useful for Accessible Web DevelopmentPeya Mowar, Yi-Hao Peng, Jason Wu, Aaron Steinfeld 等CHI 2025 · 被引用 22 次
- Memory Reviver: Supporting Photo-Collection Reminiscence for People with Visual Impairment via a Proactive ChatbotShuchang Xu, Chang Chen, Zichen Liu, Xiaofu Jin 等UIST 2024 · 被引用 21 次
- OutlineSpark: Igniting AI-powered Presentation Slides Creation from Computational Notebooks through OutlinesFengjie Wang, Yanna Lin, Leni Yang, Haotian Li 等CHI 2024 · 被引用 19 次
- The Sky is the Limit: Understanding How Generative AI can Enhance Screen Reader Users' Experience with Productivity ApplicationsMinoli Perera, Swamy Ananthanarayan, Cagatay Goncu, Kim MarriottCHI 2025 · 被引用 12 次
它引用的顶会 Paper13
- DiT: Self-supervised Pre-training for Document Image TransformerJunlong Li, Yiheng Xu, Tengchao Lv, Lei Cui 等ACM MM 2022 · 被引用 184 次
- Twitter A11y: A Browser Extension to Make Twitter Images AccessibleCole Gleason, Amy Pavel, Emma McCamey, Christina Low 等CHI 2020 · 被引用 123 次
- Rescribe: Authoring and Automatically Editing Audio DescriptionsAmy Pavel, Gabriel Reyes, Jeffrey P. BighamUIST 2020 · 被引用 72 次
- Automatic Generation of Two-Level Hierarchical Tutorials from Instructional Makeup VideosAnh Truong, Peggy Chi, David Salesin, Irfan Essa 等CHI 2021 · 被引用 57 次
- ImageExplorer: Multi-Layered Touch Exploration to Encourage Skepticism Towards Imperfect AI-Generated Image CaptionsJaewook Lee, Jaylin Herskovitz, Yi-Hao Peng, Anhong GuoCHI 2022 · 被引用 55 次
相关 Paper
- Understanding Blind Screen-Reader Users' Experiences of Digital ArtboardsAnastasia Schaadhardt, Alexis Hiniker, Jacob O. WobbrockCHI 2021 · 被引用 44 次
- Psychologically-inspired, unsupervised inference of perceptual groups of GUI widgets from GUI imagesMulong Xie, Zhenchang Xing, Sidong Feng, Xiwei Xu 等FSE 2022 · 被引用 30 次
- Diffscriber: Describing Visual Design Changes to Support Mixed-Ability Collaborative Presentation AuthoringYi-Hao Peng, Jason Wu, Jeffrey P. Bigham, Amy PavelUIST 2022 · 被引用 26 次
- Say It All: Feedback for Improving Non-Visual Presentation AccessibilityYi-Hao Peng, JiWoong Jang, Jeffrey P. Bigham, Amy PavelCHI 2021 · 被引用 46 次
- Enhancing Revisitation in Touchscreen Reading for Visually Impaired People with Semantic Navigation DesignZhichun Li, Yu Jiang, Xiaochen Liu, Yuhang Zhao 等UbiComp 2022 · 被引用 7 次
