ProgramAlly: Creating Custom Visual Access Programs via Multi-Modal End-User Programming
Jaylin Herskovitz, Andi Xu, Rahaf Alharbi, Anhong Guo
摘要
Existing visual assistive technologies are built for simple and common use cases, and have few avenues for blind people to customize their functionalities. Drawing from prior work on DIY assistive technology, this paper investigates end-user programming as a means for users to create and customize visual access programs to meet their unique needs. We introduce ProgramAlly, a system for creating custom filters for visual information, e.g., ‘find NUMBER on BUS’, leveraging three end-user programming approaches: block programming, natural language, and programming by example. To implement ProgramAlly, we designed a representation of visual filtering tasks based on scenarios encountered by blind people, and integrated a set of on-device and cloud models for generating and running these programs. In user studies with 12 blind adults, we found that participants preferred different programming modalities depending on the task, and envisioned using visual access programs to address unique accessibility challenges that are otherwise difficult with existing applications. Through ProgramAlly, we present an exploration of how blind end-users can create visual access programs to customize and control their experiences.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- "It's trained by non-disabled people": Evaluating How Image Quality Affects Product Captioning with Vision-Language ModelsKapil Garg, Xinru Tang, Jimin Heo, Dwayne R. Morgan 等CHI 2026 · 被引用 2 次
- Not Seeing the Whole Picture: Challenges and Opportunities in Using AI for Co-Making Physical, DIY-AT for People with Visual ImpairmentsBen Kosa, Hsuanling Lee, Jasmine Li, Sanbrita Mondal 等CHI 2026 · 被引用 1 次
它引用的顶会 Paper5
- Mapping Natural Language Instructions to Mobile UI Action SequencesYang Li, Jiacong He, Xin Zhou, Yuan Zhang 等ACL 2020 · 被引用 75 次
- Hacking, Switching, Combining: Understanding and Supporting DIY Assistive Technology Design by Blind PeopleJaylin Herskovitz, Andi Xu, Rahaf Alharbi, Anhong GuoCHI 2023 · 被引用 53 次
- PSST: Enabling Blind or Visually Impaired Developers to Author Sonifications of Streaming Sensor DataVenkatesh Potluri, John Thompson, James Devine, Bongshin Lee 等UIST 2022 · 被引用 25 次
- Explaining CLIP's Performance Disparities on Data from Blind/Low Vision UsersDaniela Massiceti, Camilla Longden, Agnieszka Slowik, Samuel Wills 等CVPR 2024 · 被引用 3 次
- YOLO-World: Real-Time Open-Vocabulary Object DetectionTianheng Cheng, Lin Song, Yixiao Ge, Wenyu Liu 等CVPR 2024
相关 Paper
- Do-It-Yourself AAC: Co-Designing User-Programmable AI Communication Tools with People with AphasiaJong Ho Lee, Stephanie ValenciaCHI 2026 · 被引用 1 次
- Lost in Instructions: Study of Blind Users' Experiences with DIY Manuals and AI-Rewritten Instructions for Assembly, Operation, and Troubleshooting of Tangible ProductsMonalika Padma Reddy, Aruna Balasubramanian, Jiawei Zhou, Xiaojun Bi 等CHI 2026 · 被引用 1 次
- Say It My Way: Exploring Control in Conversational Visual Question Answering with Blind UsersFarnaz Zamiri Zeraati, Yang Trista Cao, Yuehan Qiao, Hal Daumé III 等CHI 2026 · 被引用 1 次
- Towards LLM-powered Assistive Drone for Blind and Low Vision UsersYize Wei, Ibnu Taimiyyah Bin Adam, Hanjun Wu, Moritz Messerschmidt 等CHI 2026 · 被引用 1 次
- How Multimodal Large Language Models Support Access to Visual Information: A Diary Study With Blind and Low Vision PeopleRicardo E. Gonzalez Penuela, Crescentia Jung, Sharon Y. Lin, Ruiying Hu 等CHI 2026 · 被引用 1 次
