Decompose More and Aggregate Better: Two Closer Looks at Frequency Representation Learning for Human Motion Prediction
Xuehao Gao, Shaoyi Du, Yang Wu, Yang Yang
摘要
Encouraged by the effectiveness of encoding temporal dynamics within the frequency domain, recent human motion prediction systems prefer to first convert the motion representation from the original pose space into the frequency space. In this paper, we introduce two closer looks at effective frequency representation learning for robust motion prediction and summarize them as: decompose more and aggregate better. Motivated by these two insights, we develop two powerful units that factorize the frequency representation learning task with a novel decompositionaggregation two-stage strategy: (1) frequency decomposition unit unweaves multi-view frequency representations from an input body motion by embedding its frequency features into multiple spaces; (2) feature aggregation unit deploys a series of intra-space and inter-space feature aggregation layers to collect comprehensive frequency representations from these spaces for robust human motion prediction. As evaluated on large-scale datasets, we develop a strong baseline model for the human motion prediction task that outperforms state-of-the-art methods by large margins: 8%∼12% on Human3.6M, 3%∼7% on CMU MoCap, and 7%∼10% on 3DPW.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- Frequency Guidance Matters: Skeletal Action Recognition by Frequency-Aware Mixed TransformerWenhan Wu, Ce Zheng, Zihao Yang, Chen Chen 等ACM MM 2024 · 被引用 16 次
- Existence Is Chaos: Enhancing 3D Human Motion Prediction with Uncertainty ConsiderationZhihao Wang, Yulin Zhou, Ningyu Zhang, Xiaosong Yang 等AAAI 2024 · 被引用 7 次
- Frequency-Semantic Enhanced Variational Autoencoder for Zero-Shot Skeleton-Based Action RecognitionWenhan Wu, Zhishuai Guo, Chen Chen, Hongfei Xue 等ICCV 2025 · 被引用 4 次
- HVIS: A Human-like Vision and Inference System for Human Motion PredictionKedi Lyu, Haipeng Chen, Zhenguang Liu, Yifang Yin 等AAAI 2025 · 被引用 3 次
- Towards Practical Human Motion Prediction with LiDAR Point CloudsXiao Han, Yiming Ren, Yichen Yao, Yujing Sun 等ACM MM 2024 · 被引用 2 次
它引用的顶会 Paper11
- Learning Trajectory Dependencies for Human Motion PredictionWei Mao, Miaomiao Liu, Mathieu Salzmann, Hongdong LiICCV 2019 · 被引用 534 次
- MSR-GCN: Multi-Scale Residual Graph Convolution Networks for Human Motion PredictionLingwei Dang, Yongwei Nie, Chengjiang Long, Qing Zhang 等ICCV 2021 · 被引用 252 次
- Space-Time-Separable Graph Convolutional Network for Pose ForecastingTheodoros Sofianos, Alessio Sampieri, Luca Franco, Fabio GalassoICCV 2021 · 被引用 188 次
- Progressively Generating Better Initial Guesses Towards Next Stages for High-Quality Human Motion PredictionTiezheng Ma, Yongwei Nie, Chengjiang Long, Qing Zhang 等CVPR 2022 · 被引用 150 次
- Multi-Person 3D Motion Prediction with Multi-Range TransformersJiashun Wang, Huazhe Xu, Medhini Narasimhan, Xiaolong WangNeurIPS 2021 · 被引用 102 次
相关 Paper
- HUMOF: Human Motion Forecasting in Interactive Social ScenesCaiyi Sun, Yujing Sun, Xiao Han, Zemin Yang 等ICLR 2026 · 被引用 2 次
- Uncertainty-Aware Human Mesh Recovery from Video by Learning Part-Based 3D DynamicsGun-Hee Lee, Seong-Whan LeeICCV 2021 · 被引用 30 次
- Dynamic Multiscale Graph Neural Networks for 3D Skeleton Based Human Motion PredictionMaosen Li, Siheng Chen, Yangheng Zhao, Ya Zhang 等CVPR 2020
- Spatio-Temporal Branching for Motion Prediction using Motion IncrementsJiexin Wang, Yujie Zhou, Wenwen Qiang, Ying Ba 等ACM MM 2023 · 被引用 11 次
- Breaking the Passive Learning Trap: An Active Perception Strategy for Human Motion PredictionJuncheng Hu, Zijian Zhang, Zeyu Wang, Guoyu Wang 等AAAI 2026
