CookAR: Affordance Augmentations in Wearable AR to Support Kitchen Tool Interactions for People with Low Vision
Jaewook Lee, Andrew D. Tjahjadi, Jiho Kim, Junpu Yu, Minji Park, Jiawen Zhang, Jon E. Froehlich, Yapeng Tian, Yuhang Zhao
Abstract
Cooking is a central activity of daily living, supporting independence as well as mental and physical health. However, prior work has highlighted key barriers for people with low vision (LV) to cook, particularly around safely interacting with tools, such as sharp knives or hot pans. Drawing on recent advancements in computer vision (CV), we present CookAR, a head-mounted AR system with real-time object affordance augmentations to support safe and efficient interactions with kitchen tools. To design and implement CookAR, we collected and annotated the first egocentric dataset of kitchen tool affordances, fine-tuned an affordance segmentation model, and developed an AR system with a stereo camera to generate visual augmentations. To validate CookAR, we conducted a technical evaluation of our fine-tuned model as well as a qualitative lab study with 10 LV participants for suitable augmentation design. Our technical evaluation demonstrates that our model outperforms the baseline on our tool affordance dataset, while our user study indicates a preference for affordance augmentations over the traditional whole object augmentations.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0db37ff3-1a2f-407a-882d-910196be9190Cited by top-tier papers10
- Sensible Agent: A Framework for Unobtrusive Interaction with Proactive AR AgentsGeonsun Lee, Min Xia, Nels Numan, Xun Qian et al.UIST 2025 · 15 citations
- Seeing with the Hands: A Sensory Substitution That Supports Manual InteractionsShan-Yuan Teng, Gene S.-H. Kim, Xuanyou Liu, Pedro LopesCHI 2025 · 13 citations
- Vid2Coach: Transforming How-To Videos into Task AssistantsMina Huh, Zihui Xue, Ujjaini Das, Kumar Ashutosh et al.UIST 2025 · 9 citations
- VisiMark: Characterizing and Augmenting Landmarks for People with Low Vision in Augmented Reality to Support Indoor NavigationRuijia Chen, Junru Jiang, Pragati Maheshwary, Brianna R. Cochran et al.CHI 2025 · 7 citations
- StreetViewAI: Making Street View Accessible Using Context-Aware Multimodal AIJon E. Froehlich, Alexander J. Fiannaca, Nimer Jaber, Victor Tsaran et al.UIST 2025 · 5 citations
Builds on6
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- Segment AnythingAlexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao et al.ICCV 2023 · 13,211 citations
- The Effectiveness of Visual and Audio Wayfinding Guidance on Smartglasses for People with Low VisionYuhang Zhao, Elizabeth Kupferstein, Hathaitorn Rojnirun, Leah Findlater et al.CHI 2020 · 89 citations
- Wearable Subtitles: Augmenting Spoken Communication with Lightweight Eyewear for All-day CaptioningAlex Olwal, Kevin Balke, Dmitrii Votintcev, Thad Starner et al.UIST 2020 · 62 citations
- Learning Affordance Grounding from Exocentric ImagesHongchen Luo, Wei Zhai, Jing Zhang, Yang Cao et al.CVPR 2022 · 49 citations
Related papers
- Multi-label affordance mapping from egocentric visionLorenzo Mur-Labadia, Josechu J. Guerrero, Ruben Martinez-CantinICCV 2023 · 26 citations
- AROMA: Mixed-Initiative AI Assistance for Non-Visual Cooking by Grounding Multimodal Information Between Reality and VideosZheng Ning, Leyang Li, Daniel Killough, JooYoung Seo et al.UIST 2025 · 1 citation
- Enhancing Obstacle Visibility with Augmented Reality Improves Mobility in People with Low VisionLior Maman, Ilan Vol, Sarit Felicia Anais SzpiroIEEE VR 2025 · 3 citations
- RASSAR: Room Accessibility and Safety Scanning in Augmented RealityXia Su, Han Zhang, Kaiming Cheng, Jaewook Lee et al.CHI 2024 · 20 citations
- Satori 悟り: Towards Proactive AR Assistant with Belief-Desire-Intention User ModelingChenyi Li, Guande Wu, Gromit Yeuk-Yin Chan, Dishita G. Turakhia et al.CHI 2025 · 49 citations
