OCHID-Fi: Occlusion-Robust Hand Pose Estimation in 3D via RF-Vision
Shujie Zhang, Tianyue Zheng, Zhe Chen, Jingzhi Hu, Abdelwahed Khamis, Jiajun Liu, Jun Luo
摘要
Hand Pose Estimation (HPE) is crucial to many applications, but conventional cameras-based CM-HPE methods are completely subject to Line-of-Sight (LoS), as cameras cannot capture occluded objects. In this paper, we propose to exploit Radio-Frequency-Vision (RF-vision) capable of bypassing obstacles for achieving occluded HPE, and we introduce OCHID-Fi as the first RF-HPE method with 3D pose estimation capability. OCHID-Fi employs wideband RF sensors widely available on smart devices (e.g., iPhones) to probe 3D human hand pose and extract their skeletons behind obstacles. To overcome the challenge in labeling RF imaging given its human incomprehensible nature, OCHID-Fi employs a cross-modality and cross-domain training process. It uses a pre-trained CM-HPE network and a synchronized CM/RF dataset, to guide the training of its complex-valued RF-HPE network under LoS conditions. It further transfers knowledge learned from labeled LoS domain to unlabeled occluded domain via adversarial learning, enabling OCHID-Fi to generalize to unseen occluded scenarios. Experimental results demonstrate the superiority of OCHID-Fi: it achieves comparable accuracy to CM-HPE under normal conditions while maintaining such accuracy even in occluded scenarios, with empirical evidence for its generalizability to new domains.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper15
- Channel Augmented Joint Learning for Visible-Infrared RecognitionMang Ye, Weijian Ruan, Bo Du, Mike Zheng ShouICCV 2021 · 被引用 310 次
- Occlusion Robust Face Recognition Based on Mask Learning With Pairwise Differential Siamese NetworkLingxue Song, Dihong Gong, Zhifeng Li, Changsong Liu 等ICCV 2019 · 被引用 225 次
- Person-in-WiFi: Fine-Grained Person Perception Using WiFiFei Wang, Sanping Zhou, Stanislav Panev, Jinsong Han 等ICCV 2019 · 被引用 199 次
- Distilling Knowledge From a Deep Pose Regressor NetworkMuhamad Risqi Utama Saputra, Pedro Porto Buarque de Gusmão, Yasin Almalioglu, Andrew Markham 等ICCV 2019 · 被引用 116 次
- Multi-View Radar Semantic SegmentationArthur Ouaknine, Alasdair Newson, Patrick Pérez, Florence Tupin 等ICCV 2021 · 被引用 98 次
相关 Paper
- RF-URL: unsupervised representation learning for RF sensingRuiyuan Song, Dongheng Zhang, Zhi Wu, Cong Yu 等MobiCom 2022 · 被引用 62 次
- Through-Wall Human Mesh Recovery Using Radio SignalsMingmin Zhao, Yingcheng Liu, Aniruddh Raghu, Hang Zhao 等ICCV 2019 · 被引用 127 次
- RF-HOI: Recognize Human-Object Interaction with Radio Frequency SignalsLihao Wang, Linlu Gao, Jiacan Yu, Yanyu Lin 等UbiComp 2026
- Learning Longterm Representations for Person Re-Identification Using Radio SignalsLijie Fan, Tianhong Li, Rongyao Fang, Rumen Hristov 等CVPR 2020
- Vision Meets Wireless Positioning: Effective Person Re-identification with Recurrent Context PropagationYiheng Liu, Wengang Zhou, Mao Xi, Sanjing Shen 等ACM MM 2020 · 被引用 9 次
