Generative Multi-View Human Action Recognition
Lichen Wang, Zhengming Ding, Zhiqiang Tao, Yunyu Liu, Yun Fu
Abstract
Multi-view action recognition targets to integrate complementary information from different views to improve classification performance. It is a challenging task due to the distinct gap between heterogeneous feature domains. Moreover, most existing methods neglect to consider the incomplete multi-view data, which limits their potential compatibility in real-world applications. In this work, we propose a Generative Multi-View Action Recognition (GM-VAR) framework to address the challenges above. The adversarial generative network is leveraged to generate one view conditioning on the other view, which fully explores the latent connections in both intra-view and cross-view aspects. Our approach enhances the model robustness by employing adversarial training, and naturally handles the incomplete view case by imputing the missing data. Moreover, an effective View Correlation Discovery Network (VCDN) is proposed to further fuse the multi-view information in a higher-level label space. Extensive experiments demonstrate the effectiveness of our proposed approach by comparing with state-of-the-art algorithms 1 .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 69eee172-d15f-4036-88b6-d816e811c18bCited by top-tier papers7
- UniFi: A Unified Framework for Generalizable Gesture Recognition with Wi-Fi Signals Using Consistency-guided Multi-View NetworksYan Liu, Anlan Yu, Leye Wang, Bin Guo et al.UbiComp 2024 · 57 citations
- Multi-Level Confidence Learning for Trustworthy Multimodal ClassificationXiao Zheng, Chang Tang, Zhiguo Wan, Chengyu Hu et al.AAAI 2023 · 41 citations
- DVANet: Disentangling View and Action Features for Multi-View Action RecognitionNyle Siddiqui, Praveen Tirupattur, Mubarak ShahAAAI 2024 · 39 citations
- Weakly-Supervised Online Action Segmentation in Multi-View Instructional VideosReza Ghoddoosian, Isht Dwivedi, Nakul Agarwal, Chiho Choi et al.CVPR 2022 · 22 citations
- Generative Partial Visual-Tactile Fused Object ClusteringTao Zhang, Yang Cong, Gan Sun, Jiahua Dong et al.AAAI 2021 · 16 citations
Related papers
- View-normalized Skeleton Generation for Action RecognitionQingzhe Pan, Zhifu Zhao, Xuemei Xie, Jianan Li et al.ACM MM 2021 · 12 citations
- Deep Adversarial Completion for Sparse Heterogeneous Information Network EmbeddingKai Zhao, Ting Bai, Bin Wu, Bai Wang et al.WWW 2020 · 35 citations
- Attention-Induced Embedding Imputation for Incomplete Multi-View Partial Multi-Label ClassificationChengliang Liu, Jinlong Jia, Jie Wen, Yabo Liu et al.AAAI 2024 · 39 citations
- Jointly Imputing Multi-View Data with Optimal TransportYangyang Wu, Xiaoye Miao, Xinyu Huang, Jianwei YinAAAI 2023 · 13 citations
- Beyond Independence: Learning Correlated Views for Variational Incomplete Multi-View ClusteringZheming Xu, Aiyue Tang, Shidi Chen, Xuechao Zou et al.ICML 2026
