UniFi: A Unified Framework for Generalizable Gesture Recognition with Wi-Fi Signals Using Consistency-guided Multi-View Networks
Yan Liu, Anlan Yu, Leye Wang, Bin Guo, Yang Li, Enze Yi, Daqing Zhang
Abstract
In recent years, considerable endeavors have been devoted to exploring Wi-Fi-based sensing technologies by modeling the intricate mapping between received signals and corresponding human activities. However, the inherent complexity of Wi-Fi signals poses significant challenges for practical applications due to their pronounced susceptibility to deployment environments. To address this challenge, we delve into the distinctive characteristics of Wi-Fi signals and distill three pivotal factors that can be leveraged to enhance generalization capabilities of deep learning-based Wi-Fi sensing models: 1) effectively capture valuable input to mitigate the adverse impact of noisy measurements; 2) adaptively fuse complementary information from multiple Wi-Fi devices to boost the distinguishability of signal patterns associated with different activities; 3) extract generalizable features that can overcome the inconsistent representations of activities under different environmental conditions (e.g., locations, orientations). Leveraging these insights, we design a novel and unified sensing framework based on Wi-Fi signals, dubbed UniFi, and use gesture recognition as an application to demonstrate its effectiveness. UniFi achieves robust and generalizable gesture recognition in real-world scenarios by extracting discriminative and consistent features unrelated to environmental factors from pre-denoised signals collected by multiple transceivers. To achieve this, we first introduce an effective signal preprocessing approach that captures the applicable input data from noisy received signals for the deep learning model. Second, we propose a multi-view deep network based on spatio-temporal cross-view attention that integrates multi-carrier and multi-device signals to extract distinguishable information. Finally, we present the mutual information maximization as a regularizer to learn environment-invariant representations via contrastive loss without requiring access to any signals from unseen environments for practical adaptation. Extensive experiments on the Widar 3.0 dataset demonstrate that our proposed framework significantly outperforms state-of-the-art approaches in different settings (99% and 90%-98% accuracy for in-domain and cross-domain recognition without additional data collection and model training).
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 668cc88b-1e38-4da7-89ff-0c3c000a2abfCited by top-tier papers4
- SeRadar: Embracing Secondary Reflections for Human Sensing with mmWave RadarDanei Gong, Naiyu Zheng, Binbin Xie, Jie Xiong et al.MobiCom 2025 · 3 citations
- TAPOR: 3D Hand Pose Reconstruction with Fully Passive Thermal Sensing for Around-Device InteractionsXie Zhang, Chengxiao Li, Chenshu WuUbiComp 2025 · 3 citations
- UNI-FI: Integrated Multi-Task Wi-Fi SensingMengning Li, Wenye WangINFOCOM 2026 · 1 citation
- Beyond Physical Labels: Redefining Domains for Robust WiFi-based Gesture RecognitionXiang Zhang, Huan Yan, Jinyang Huang, Bin Liu et al.UbiComp 2026 · 1 citation
Builds on11
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Two-Stream Convolution Augmented Transformer for Human Activity RecognitionBing Li, Wei Cui, Wei Wang, Le Zhang et al.AAAI 2021 · 233 citations
- Model-Based Domain GeneralizationAlexander Robey, George J. Pappas, Hamed HassaniNeurIPS 2021 · 167 citations
- Towards Position-Independent Sensing for Gesture Recognition with Wi-FiRuiyang Gao, Mi Zhang, Jie Zhang, Yang Li et al.UbiComp 2021 · 143 citations
- PCL: Proxy-based Contrastive Learning for Domain GeneralizationXufeng Yao, Yang Bai, Xinyun Zhang, Yuechen Zhang et al.CVPR 2022 · 127 citations
Related papers
- Towards Robust Gesture Recognition by Characterizing the Sensing Quality of WiFi SignalsRuiyang Gao, Wenwei Li, Yaxiong Xie, Enze Yi et al.UbiComp 2022 · 82 citations
- WiAdv: Practical and Robust Adversarial Attack against WiFi-based Gesture Recognition SystemYuxuan Zhou, Huangxun Chen, Chenyu Huang, Qian ZhangUbiComp 2022 · 31 citations
- RF-URL: unsupervised representation learning for RF sensingRuiyuan Song, Dongheng Zhang, Zhi Wu, Cong Yu et al.MobiCom 2022 · 62 citations
- Wi-Learner: Towards One-shot Learning for Cross-Domain Wi-Fi based Gesture RecognitionChao Feng, Nan Wang, Yicheng Jiang, Xia Zheng et al.UbiComp 2022 · 48 citations
- CrossGR: Accurate and Low-cost Cross-target Gesture Recognition Using Wi-FiXinyi Li, Liqiong Chang, Fangfang Song, Ju Wang et al.UbiComp 2021 · 51 citations
