CookAR: Affordance Augmentations in Wearable AR to Support Kitchen Tool Interactions for People with Low Vision
Jaewook Lee, Andrew D. Tjahjadi, Jiho Kim, Junpu Yu, Minji Park, Jiawen Zhang, Jon E. Froehlich, Yapeng Tian, Yuhang Zhao
摘要
Cooking is a central activity of daily living, supporting independence as well as mental and physical health. However, prior work has highlighted key barriers for people with low vision (LV) to cook, particularly around safely interacting with tools, such as sharp knives or hot pans. Drawing on recent advancements in computer vision (CV), we present CookAR, a head-mounted AR system with real-time object affordance augmentations to support safe and efficient interactions with kitchen tools. To design and implement CookAR, we collected and annotated the first egocentric dataset of kitchen tool affordances, fine-tuned an affordance segmentation model, and developed an AR system with a stereo camera to generate visual augmentations. To validate CookAR, we conducted a technical evaluation of our fine-tuned model as well as a qualitative lab study with 10 LV participants for suitable augmentation design. Our technical evaluation demonstrates that our model outperforms the baseline on our tool affordance dataset, while our user study indicates a preference for affordance augmentations over the traditional whole object augmentations.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- Sensible Agent: A Framework for Unobtrusive Interaction with Proactive AR AgentsGeonsun Lee, Min Xia, Nels Numan, Xun Qian 等UIST 2025 · 被引用 15 次
- Seeing with the Hands: A Sensory Substitution That Supports Manual InteractionsShan-Yuan Teng, Gene S.-H. Kim, Xuanyou Liu, Pedro LopesCHI 2025 · 被引用 13 次
- Vid2Coach: Transforming How-To Videos into Task AssistantsMina Huh, Zihui Xue, Ujjaini Das, Kumar Ashutosh 等UIST 2025 · 被引用 9 次
- VisiMark: Characterizing and Augmenting Landmarks for People with Low Vision in Augmented Reality to Support Indoor NavigationRuijia Chen, Junru Jiang, Pragati Maheshwary, Brianna R. Cochran 等CHI 2025 · 被引用 7 次
- StreetViewAI: Making Street View Accessible Using Context-Aware Multimodal AIJon E. Froehlich, Alexander J. Fiannaca, Nimer Jaber, Victor Tsaran 等UIST 2025 · 被引用 5 次
它引用的顶会 Paper6
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- Segment AnythingAlexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao 等ICCV 2023 · 被引用 13,211 次
- The Effectiveness of Visual and Audio Wayfinding Guidance on Smartglasses for People with Low VisionYuhang Zhao, Elizabeth Kupferstein, Hathaitorn Rojnirun, Leah Findlater 等CHI 2020 · 被引用 89 次
- Wearable Subtitles: Augmenting Spoken Communication with Lightweight Eyewear for All-day CaptioningAlex Olwal, Kevin Balke, Dmitrii Votintcev, Thad Starner 等UIST 2020 · 被引用 62 次
- Learning Affordance Grounding from Exocentric ImagesHongchen Luo, Wei Zhai, Jing Zhang, Yang Cao 等CVPR 2022 · 被引用 49 次
相关 Paper
- Multi-label affordance mapping from egocentric visionLorenzo Mur-Labadia, Josechu J. Guerrero, Ruben Martinez-CantinICCV 2023 · 被引用 26 次
- AROMA: Mixed-Initiative AI Assistance for Non-Visual Cooking by Grounding Multimodal Information Between Reality and VideosZheng Ning, Leyang Li, Daniel Killough, JooYoung Seo 等UIST 2025 · 被引用 1 次
- Enhancing Obstacle Visibility with Augmented Reality Improves Mobility in People with Low VisionLior Maman, Ilan Vol, Sarit Felicia Anais SzpiroIEEE VR 2025 · 被引用 3 次
- RASSAR: Room Accessibility and Safety Scanning in Augmented RealityXia Su, Han Zhang, Kaiming Cheng, Jaewook Lee 等CHI 2024 · 被引用 20 次
- Satori 悟り: Towards Proactive AR Assistant with Belief-Desire-Intention User ModelingChenyi Li, Guande Wu, Gromit Yeuk-Yin Chan, Dishita G. Turakhia 等CHI 2025 · 被引用 49 次
