RAVEN: Realtime Accessibility in Virtual ENvironments for Blind and Low-Vision People
Xinyun Cao, Kexin Phyllis Ju, Chenglin Li, Venkatesh Potluri, Dhruv Jain
摘要
As virtual 3D environments become more prevalent, equitable access is essential for blind and low-vision (BLV) users, who face challenges with spatial awareness, navigation, and interaction. Prior work has explored supplementing visual information with auditory or haptic modalities, but these methods are static and offer limited support for dynamic, in-context adaptation. Recent advances in generative AI allow users to query and modify 3D scenes via natural language, introducing a paradigm that offers greater flexibility and control for accessibility. We present RAVEN, a system that enables BLV users to issue queries and modification prompts to improve the runtime accessibility of 3D virtual scenes. We evaluated RAVEN with eight BLV people and six Unity developers, generating empirical insights into how conversational programming can support personalized accessibility in 3D environments. Our work highlights both the promise of natural language interaction—intuitive, flexible, and empowering—and the challenges of ensuring reliability, transparency, and trust in generative AI–driven accessibility systems.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper22
- Audio Flamingo: A Novel Audio Language Model with Few-Shot Learning and Dialogue AbilitiesZhifeng Kong, Arushi Goel, Rohan Badlani, Wei Ping 等ICML 2024 · 被引用 207 次
- LLMR: Real-time Prompting of Interactive Worlds using Large Language ModelsFernanda De La Torre, Cathy Mengying Fang, Han Huang, Andrzej Banburski-Fahey 等CHI 2024 · 被引用 124 次
- Should we be going MAD? A Look at Multi-Agent Debate Strategies for LLMsAndries P. Smit, Nathan Grinsztajn, Paul Duckworth, Thomas D. Barrett 等ICML 2024 · 被引用 82 次
- GenAssist: Making Image Generation AccessibleMina Huh, Yi-Hao Peng, Amy PavelUIST 2023 · 被引用 58 次
- "They only care to show us the wheelchair": disability representation in text-to-image AI modelsKelly Avery Mack, Rida Qadri, Remi Denton, Shaun K. Kane 等CHI 2024 · 被引用 57 次
相关 Paper
- Understanding the Use of a Large Language Model-Powered Guide to Make Virtual Reality Accessible for Blind and Low Vision PeopleJazmin Collins, Sharon Y. Lin, Tianqi Liu, Andrea Stevenson Won 等CHI 2026 · 被引用 3 次
- "Game Changer" or "Overenthusiastic Drunk Acquaintance"? Generative AI Use by Blind and Low Vision Software Professionals in the WorkplaceYoonha Cha, Victoria Jackson, Lauren Shu, Stacy Branham 等ICSE 2026
- Say It My Way: Exploring Control in Conversational Visual Question Answering with Blind UsersFarnaz Zamiri Zeraati, Yang Trista Cao, Yuehan Qiao, Hal Daumé III 等CHI 2026 · 被引用 1 次
- The Sky is the Limit: Understanding How Generative AI can Enhance Screen Reader Users' Experience with Productivity ApplicationsMinoli Perera, Swamy Ananthanarayan, Cagatay Goncu, Kim MarriottCHI 2025 · 被引用 12 次
- "Hey Model!" - Natural User Interactions and Agency in Accessible Interactive 3D ModelsSamuel Reinders, Matthew Butler, Kim MarriottCHI 2020 · 被引用 22 次
