CAFI-AR: Contact-aware Freehand Interaction with AR Objects
Xiao Tang, Ruihui Li, Chi-Wing Fu
Abstract
Freehand interaction enhances user experience, allowing one to use bare hands to manipulate virtual objects in AR. Yet, it remains challenging to accurately and efficiently detect contacts between real hand and virtual object, due to the imprecise captured/estimated hand geometry. This paper presents CAFI-AR, a new approach for Contact-Aware Freehand Interaction with virtual AR objects, enabling us to automatically detect hand-object contacts in real-time with low latency. Specifically, we formulate a compact deep architecture to efficiently learn to predict hand action and contact moment from sequences of captured RGB images relative to the 3D virtual object. To train the architecture for detecting contacts on AR objects, we build a new dataset with 4,008 frame sequences, each with annotated hand-object interaction information. Further, we integrate CAFI-AR into our prototyping AR system and develop various interactive scenarios, demonstrating fine-grained contact-aware interactions on a rich variety of virtual AR objects, which cannot be achieved by existing AR interaction approaches. Lastly, we also evaluate CAFI-AR, quantitatively and qualitatively, through two user studies to demonstrate its effectiveness in terms of accurately detecting the hand-object contacts and promoting fluid freehand interactions
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get eb3b1b5b-1306-490e-9bf7-1a49f0d98e72Related papers
- GrabAR: Occlusion-aware Grabbing Virtual Objects in ARXiao Tang, Xiaowei Hu, Chi-Wing Fu, Daniel Cohen-OrUIST 2020 · 29 citations
- ARnnotate: An Augmented Reality Interface for Collecting Custom Dataset of 3D Hand-Object Interaction Pose EstimationXun Qian, Fengming He, Xiyun Hu, Tianyi Wang et al.UIST 2022 · 16 citations
- Real-Time Multimodal Fingertip Contact Detection via Depth and Motion Fusion for Vision-Based Human–Computer InteractionMukhiddin Toshpulatov, Wookey Lee, Suan Lee, Geehyuk LeeCVPR 2026
- MEgATrack: monochrome egocentric articulated hand-tracking for virtual realityShangchen Han, Beibei Liu, Randi Cabezas, Christopher D. Twigg et al.SIGGRAPH 2020 · 207 citations
- A Multimodal Approach for Targeting Error Detection in Virtual Reality Using Implicit User BehaviorNaveen Sendhilnathan, Ting Zhang, David Bethge, Michael Nebeling et al.CHI 2025 · 3 citations
