Making the Invisible Visible: Action Recognition Through Walls and Occlusions
Tianhong Li, Lijie Fan, Mingmin Zhao, Yingcheng Liu, Dina Katabi
Abstract
Understanding people's actions and interactions typically depends on seeing them. Automating the process of action recognition from visual data has been the topic of much research in the computer vision community. But what if it is too dark, or if the person is occluded or behind a wall? In this paper, we introduce a neural network model that can detect human actions through walls and occlusions, and in poor lighting conditions. Our model takes radio frequency (RF) signals as input, generates 3D human skeletons as an intermediate representation, and recognizes actions and interactions of multiple people over time. By translating the input to an intermediate skeleton-based representation, our model can learn from both vision-based and RF-based datasets, and allow the two tasks to help each other. We show that our model achieves comparable accuracy to vision-based action recognition systems in visible scenarios, yet continues to work accurately when people are not visible, hence addressing scenarios that are beyond the limit of today's vision-based action recognition. * Indicates equal contribution. Ordering determined by inverse alphabetical order.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 94e3f93f-b27a-4eb5-b071-089cc6e2e307Cited by top-tier papers19
- Targeted Supervised Contrastive Learning for Long-Tailed RecognitionTianhong Li, Peng Cao, Yuan Yuan, Lijie Fan et al.CVPR 2022 · 196 citations
- Through-Wall Human Mesh Recovery Using Radio SignalsMingmin Zhao, Yingcheng Liu, Aniruddh Raghu, Hang Zhao et al.ICCV 2019 · 127 citations
- SpiroSonic: monitoring human lung function via acoustic sensing on commodity smartphonesXingzhe Song, Boyuan Yang, Ge Yang, Ruirong Chen et al.MobiCom 2020 · 97 citations
- Motion Prediction using Trajectory CuesZhenguang Liu, Pengxiang Su, Shuang Wu, Xuanjing Shen et al.ICCV 2021 · 63 citations
- mmBody Benchmark: 3D Body Reconstruction Dataset and Analysis for Millimeter Wave RadarAnjun Chen, Xiangyu Wang, Shaohao Zhu, Yanxu Li et al.ACM MM 2022 · 63 citations
Builds on1
Related papers
- XRF55: A Radio Frequency Dataset for Human Indoor Action AnalysisFei Wang, Yizhe Lv, Mengdie Zhu, Han Ding et al.UbiComp 2024 · 47 citations
- SkeleTR: Towards Skeleton-based Action Recognition in the WildHaodong Duan, Mingze Xu, Bing Shuai, Davide Modolo et al.ICCV 2023 · 38 citations
- OccMesh: Occlusion-aware Multi-user 3D Human Mesh Reconstruction Using mmWave SignalsHaoran Cao, Jiadi Yu, Hao Kong, Yi-Chao Chen et al.UbiComp 2025 · 4 citations
- OCHID-Fi: Occlusion-Robust Hand Pose Estimation in 3D via RF-VisionShujie Zhang, Tianyue Zheng, Zhe Chen, Jingzhi Hu et al.ICCV 2023 · 11 citations
- PREDICT & CLUSTER: Unsupervised Skeleton Based Action RecognitionKun Su, Xiulong Liu, Eli ShlizermanCVPR 2020
