Neural Koopman Pooling: Control-Inspired Temporal Dynamics Encoding for Skeleton-Based Action Recognition
Xinghan Wang, Xin Xu, Yadong Mu
Abstract
Skeleton-based human action recognition is becoming increasingly important in a variety of fields. Most existing works train a CNN or GCN based backbone to extract spatial-temporal features, and use temporal average/max pooling to aggregate the information. However, these pooling methods fail to capture high-order dynamics information. To address the problem, we propose a plug-andplay module called Koopman pooling, which is a parameterized high-order pooling technique based on Koopman theory. The Koopman operator linearizes a non-linear dynamics system, thus providing a way to represent the complex system through the dynamics matrix, which can be used for classification. We also propose an eigenvalue normalization method to encourage the learned dynamics to be non-decaying and stable. Besides, we also show that our Koopman pooling framework can be easily extended to one-shot action recognition when combined with Dynamic Mode Decomposition. The proposed method is evaluated on three benchmark datasets, namely NTU RGB+D 60, 120 and NW-UCLA. Our experiments clearly demonstrate that Koopman pooling significantly improves the performance under both full-dataset and one-shot settings.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext dac46688-631f-466c-a96a-e68e6118405dCited by top-tier papers5
- LLMs are Good Action RecognizersHaoxuan Qu, Yujun Cai, Jun LiuCVPR 2024 · 37 citations
- Multi-Modality Co-Learning for Efficient Skeleton-based Action RecognitionJinfu Liu, Chen Chen, Mengyuan LiuACM MM 2024 · 27 citations
- Motion Matters: Motion-guided Modulation Network for Skeleton-based Micro-Action RecognitionJihao Gu, Kun Li, Fei Wang, Yanyan Wei et al.ACM MM 2025 · 23 citations
- Frequency Guidance Matters: Skeletal Action Recognition by Frequency-Aware Mixed TransformerWenhan Wu, Ce Zheng, Zihao Yang, Chen Chen et al.ACM MM 2024 · 16 citations
- MaskCLR: Attention-Guided Contrastive Learning for Robust Action Representation LearningMohamed Abdelfattah, Mariam Hassan, Alexandre AlahiCVPR 2024
Builds on19
- Channel-wise Topology Refinement Graph Convolution for Skeleton-Based Action RecognitionYuxin Chen, Ziqi Zhang, Chunfeng Yuan, Bing Li et al.ICCV 2021 · 871 citations
- InfoGCN: Representation Learning for Human Skeleton-based Action RecognitionHyung-Gun Chi, Myoung Hoon Ha, Seung-geun Chi, Sang Wan Lee et al.CVPR 2022 · 383 citations
- Dynamic GCN: Context-enriched Topology Learning for Skeleton-based Action RecognitionFanfan Ye, Shiliang Pu, Qiaoyong Zhong, Chao Li et al.ACM MM 2020 · 348 citations
- Multi-Scale Spatial Temporal Graph Convolutional Network for Skeleton-Based Action RecognitionZhan Chen, Sicheng Li, Bing Yang, Qinghan Li et al.AAAI 2021 · 341 citations
- Hierarchically Decomposed Graph Convolutional Networks for Skeleton-Based Action RecognitionJungho Lee, Minhyeok Lee, Dogyoon Lee, Sangyoun LeeICCV 2023 · 236 citations
Related papers
- Spatio-Temporal Fusion for Human Action Recognition via Joint Trajectory GraphYaolin Zheng, Hongbo Huang, Xiuying Wang, Xiaoxu Yan et al.AAAI 2024 · 22 citations
- Efficient Parametric SVD of Koopman Operator for Stochastic Dynamical SystemsMinchan Jeong, Jongha Ryu, Se-Young Yun, Gregory W. WornellNeurIPS 2025 · 6 citations
- When Graph Neural Networks Meet Dynamic Mode DecompositionDai Shi, Lequan Lin, Andi Han, Zhiyong Wang et al.ICLR 2025
- Towards To-a-T Spatio-Temporal Focus for Skeleton-Based Action RecognitionLipeng Ke, Kuan-Chuan Peng, Siwei LyuAAAI 2022 · 47 citations
- Leveraging Spatio-Temporal Dependency for Skeleton-Based Action RecognitionJungho Lee, Minhyeok Lee, Suhwan Cho, Sungmin Woo et al.ICCV 2023 · 28 citations
