OCHID-Fi: Occlusion-Robust Hand Pose Estimation in 3D via RF-Vision
Shujie Zhang, Tianyue Zheng, Zhe Chen, Jingzhi Hu, Abdelwahed Khamis, Jiajun Liu, Jun Luo
Abstract
Hand Pose Estimation (HPE) is crucial to many applications, but conventional cameras-based CM-HPE methods are completely subject to Line-of-Sight (LoS), as cameras cannot capture occluded objects. In this paper, we propose to exploit Radio-Frequency-Vision (RF-vision) capable of bypassing obstacles for achieving occluded HPE, and we introduce OCHID-Fi as the first RF-HPE method with 3D pose estimation capability. OCHID-Fi employs wideband RF sensors widely available on smart devices (e.g., iPhones) to probe 3D human hand pose and extract their skeletons behind obstacles. To overcome the challenge in labeling RF imaging given its human incomprehensible nature, OCHID-Fi employs a cross-modality and cross-domain training process. It uses a pre-trained CM-HPE network and a synchronized CM/RF dataset, to guide the training of its complex-valued RF-HPE network under LoS conditions. It further transfers knowledge learned from labeled LoS domain to unlabeled occluded domain via adversarial learning, enabling OCHID-Fi to generalize to unseen occluded scenarios. Experimental results demonstrate the superiority of OCHID-Fi: it achieves comparable accuracy to CM-HPE under normal conditions while maintaining such accuracy even in occluded scenarios, with empirical evidence for its generalizability to new domains.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 874b4533-69fd-41e1-beb4-ba6d0e9b6141Builds on15
- Channel Augmented Joint Learning for Visible-Infrared RecognitionMang Ye, Weijian Ruan, Bo Du, Mike Zheng ShouICCV 2021 · 310 citations
- Occlusion Robust Face Recognition Based on Mask Learning With Pairwise Differential Siamese NetworkLingxue Song, Dihong Gong, Zhifeng Li, Changsong Liu et al.ICCV 2019 · 225 citations
- Person-in-WiFi: Fine-Grained Person Perception Using WiFiFei Wang, Sanping Zhou, Stanislav Panev, Jinsong Han et al.ICCV 2019 · 199 citations
- Distilling Knowledge From a Deep Pose Regressor NetworkMuhamad Risqi Utama Saputra, Pedro Porto Buarque de Gusmão, Yasin Almalioglu, Andrew Markham et al.ICCV 2019 · 116 citations
- Multi-View Radar Semantic SegmentationArthur Ouaknine, Alasdair Newson, Patrick Pérez, Florence Tupin et al.ICCV 2021 · 98 citations
Related papers
- RF-URL: unsupervised representation learning for RF sensingRuiyuan Song, Dongheng Zhang, Zhi Wu, Cong Yu et al.MobiCom 2022 · 62 citations
- Through-Wall Human Mesh Recovery Using Radio SignalsMingmin Zhao, Yingcheng Liu, Aniruddh Raghu, Hang Zhao et al.ICCV 2019 · 127 citations
- RF-HOI: Recognize Human-Object Interaction with Radio Frequency SignalsLihao Wang, Linlu Gao, Jiacan Yu, Yanyu Lin et al.UbiComp 2026
- Learning Longterm Representations for Person Re-Identification Using Radio SignalsLijie Fan, Tianhong Li, Rongyao Fang, Rumen Hristov et al.CVPR 2020
- Vision Meets Wireless Positioning: Effective Person Re-identification with Recurrent Context PropagationYiheng Liu, Wengang Zhou, Mao Xi, Sanjing Shen et al.ACM MM 2020 · 9 citations
