X-World: Accessibility, Vision, and Autonomy Meet
Jimuyang Zhang, Minglan Zheng, Matthew Boyd, Eshed Ohn-Bar
摘要
An important issue facing vision-based intelligent systems today is the lack of accessibility-aware development. A main reason for this issue is the absence of any large-scale, standardized vision benchmarks that incorporate relevant tasks and scenarios related to people with disabilities. This lack of representation hinders even preliminary analysis with respect to underlying pose, appearance, and occlusion characteristics of diverse pedestrians. What is the impact of significant occlusion from a wheelchair on instance segmentation quality? How can interaction with mobility aids, e.g., a long and narrow walking cane, be recognized robustly? To begin addressing such questions, we introduce X-World, an accessibility-centered development environment for vision-based autonomous systems. We tackle inherent data scarcity by leveraging a simulation environment to spawn dynamic agents with various mobility aids. The simulation supports generation of ample amounts of finely annotated, multi-modal data in a safe, cheap, and privacy-preserving manner. Our analysis highlights novel challenges introduced by our benchmark and tasks, as well as numerous opportunities for future developments. We further broaden our analysis using a complementary real-world evaluation benchmark of in-situ navigation by pedestrians with disabilities. Our contributions provide an initial step towards widespread deployment of vision-based agents that can perceive and model the interaction needs of diverse people with disabilities.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- XVO: Generalized Visual Odometry via Cross-Modal Self-TrainingLei Lai, Zhongkai Shangguan, Jimuyang Zhang, Eshed Ohn-BarICCV 2023 · 被引用 27 次
- SelfD: Self-Learning Large-Scale Driving Policies From the WebJimuyang Zhang, Ruizhao Zhu, Eshed Ohn-BarCVPR 2022 · 被引用 17 次
- Feedback-Guided Autonomous DrivingJimuyang Zhang, Zanming Huang, Arijit Ray, Eshed Ohn-BarCVPR 2024 · 被引用 15 次
- Motion Diversification NetworksHee Jae Kim, Eshed Ohn-BarCVPR 2024
它引用的顶会 Paper11
- Frustratingly Simple Few-Shot Object DetectionXin Wang, Thomas E. Huang, Joseph Gonzalez, Trevor Darrell 等ICML 2020 · 被引用 723 次
- Exploring the Limitations of Behavior Cloning for Autonomous DrivingFelipe Codevilla, Eder Santana, Antonio M. López, Adrien GaidonICCV 2019 · 被引用 666 次
- Meta-Learning to Detect Rare ObjectsYu-Xiong Wang, Deva Ramanan, Martial HebertICCV 2019 · 被引用 339 次
- nuScenes: A Multimodal Dataset for Autonomous DrivingHolger Caesar, Varun Bankiti, Alex H. Lang, Sourabh Vora 等CVPR 2020
- Sign Language Transformers: Joint End-to-End Sign Language Recognition and TranslationNecati Cihan Camgöz, Oscar Koller, Simon Hadfield, Richard BowdenCVPR 2020
相关 Paper
- Accessibility for Whom? Perceptions of Mobility Barriers Across Disability Groups and Implications for Designing Personalized MapsChu Li, Rock Yuren Pang, Delphine Labbé, Yochai Eisenberg 等CHI 2025 · 被引用 15 次
- R2H: Building Multimodal Navigation Helpers that Respond to Help RequestsYue Fan, Jing Gu, Kaizhi Zheng, Xin WangEMNLP 2023 · 被引用 3 次
- Accessibility Scout: Personalized Accessibility Scans of Built EnvironmentsWilliam Huang, Xia Su, Jon E. Froehlich, Yang ZhangUIST 2025 · 被引用 3 次
- Virtual Worlds Beyond Sight: Designing and Evaluating an Audio-Haptic System for Non-Visual VR ExplorationAayush Shrestha, Joseph MallochCHI 2025 · 被引用 6 次
- From Selfie Stick to Virtual Cane: Enabling Blind Exploration through Mobile Virtual RealityHao Tang, Hong Zhao, Xinpeng Liu, Zhenchao Xia 等CHI 2026
