HVIS: A Human-like Vision and Inference System for Human Motion Prediction
Kedi Lyu, Haipeng Chen, Zhenguang Liu, Yifang Yin, Yukang Lin, Yingying Jiao
Abstract
Grasping the intricacies of human motion, which involve perceiving spatio-temporal dependence and multi-scale effects, is essential for predicting human motion. While humans inherently possess the requisite skills to navigate this issue, it proves to be markedly more challenging for machines to emulate. To bridge the gap, we propose the Human-like Vision and Inference System (HVIS) for human motion prediction, which is designed to emulate human observation and forecast future movements. HVIS comprises two components: the human-like vision encode (HVE) module and the human-like motion inference (HMI) module. The HVE module mimics and refines the human visual process, incorporating a retina-analog component that captures spatiotemporal information separately to avoid unnecessary crosstalk. Additionally, a visual cortex-analogy component is designed to hierarchically extract and treat complex motion features, focusing on both global and local features of human poses. The HMI is employed to simulate the multi-stage learning model of the human brain. The spontaneous learning network simulates the neuronal fracture generation process for the adversarial generation of future motions. Subsequently, the deliberate learning network is optimized for hard-to-train joints to prevent misleading learning. Experimental results demonstrate that our method achieves new state-of-the-art performance, significantly outperforming existing methods by 19.8 % on Human3.6M, 15.7 % on CMU Mocap, and 11.1 % on G3D.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers2
- Separate to Collaborate: Dual-Stream Diffusion Model for Coordinated Piano Hand Motion SynthesisZihao Liu, Mingwen Ou, Zunnan Xu, Jiaqi Huang et al.ACM MM 2025 · 2 citations
- InterAnimate: Taming Region-Aware Diffusion Model for Realistic Human Interaction AnimationYukang Lin, Yan Hong, Zunnan Xu, Xindi Li et al.ACM MM 2025
Builds on5
- Aggregated Multi-GANs for Controlled 3D Human Motion PredictionZhenguang Liu, Kedi Lyu, Shuang Wu, Haipeng Chen et al.AAAI 2021 · 64 citations
- Locate and Verify: A Two-Stream Network for Improved Deepfake DetectionChao Shuai, Jieming Zhong, Shuang Wu, Feng Lin et al.ACM MM 2023 · 52 citations
- Motion Prediction via Joint Dependency Modeling in Phase SpacePengxiang Su, Zhenguang Liu, Shuang Wu, Lei Zhu et al.ACM MM 2021 · 22 citations
- Decompose More and Aggregate Better: Two Closer Looks at Frequency Representation Learning for Human Motion PredictionXuehao Gao, Shaoyi Du, Yang Wu, Yang YangCVPR 2023
- Dynamic Multiscale Graph Neural Networks for 3D Skeleton Based Human Motion PredictionMaosen Li, Siheng Chen, Yangheng Zhao, Ya Zhang et al.CVPR 2020
Related papers
- Learning Trajectory Dependencies for Human Motion PredictionWei Mao, Miaomiao Liu, Mathieu Salzmann, Hongdong LiICCV 2019 · 534 citations
- Structured Prediction Helps 3D Human Motion ModellingEmre Aksan, Manuel Kaufmann, Otmar HilligesICCV 2019 · 204 citations
- Fg-T2M: Fine-Grained Text-Driven Human Motion Generation via Diffusion ModelYin Wang, Zhiying Leng, Frederick W. B. Li, Shun-Cheng Wu et al.ICCV 2023 · 95 citations
- Progressively Generating Better Initial Guesses Towards Next Stages for High-Quality Human Motion PredictionTiezheng Ma, Yongwei Nie, Chengjiang Long, Qing Zhang et al.CVPR 2022 · 150 citations
- Towards Accurate 3D Human Motion Prediction From Incomplete ObservationsQiongjie Cui, Huaijiang SunCVPR 2021
