Contact-aware Human Motion Forecasting
Wei Mao, Miaomiao Liu, Richard I. Hartley, Mathieu Salzmann
摘要
In this paper, we tackle the task of scene-aware 3D human motion forecasting, which consists of predicting future human poses given a 3D scene and a past human motion. A key challenge of this task is to ensure consistency between the human and the scene, accounting for human-scene interactions. Previous attempts to do so model such interactions only implicitly, and thus tend to produce artifacts such as "ghost motion" because of the lack of explicit constraints between the local poses and the global motion. Here, by contrast, we propose to explicitly model the human-scene contacts. To this end, we introduce distance-based contact maps that capture the contact relationships between every joint and every 3D scene point at each time instant. We then develop a two-stage pipeline that first predicts the future contact maps from the past ones and the scene point cloud, and then forecasts the future human poses by conditioning them on the predicted contact maps. During training, we explicitly encourage consistency between the global motion and the local poses via a prior defined using the contact maps and future poses. Our approach outperforms the state-of-the-art human motion forecasting and human synthesis methods on both synthetic and real datasets. Our code is available at https://github.com/wei-mao-2019/ContAwareMotionPred . Recently, a few works [8, 6] have started to incorporate scene context in motion forecasting. In particular, Corona et al. [8] introduced a semantic-graph model that extracts a joint embedding of the human pose and an object of interest, such as a cup. This method, however, is ill-suited to model interactions with the whole scene itself, for example the floor or stairs that the person touches while walking. In [6], Cao et al. proposed a multi-stage pipeline that breaks down the motion forecasting into three sub-tasks: predicting a 2D goal, planning a 2D and 3D path, forecasting the 3D poses
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper19
- A Single 2D Pose with Context is Worth Hundreds for 3D Human Pose EstimationQitao Zhao, Ce Zheng, Mengyuan Liu, Chen ChenNeurIPS 2023 · 被引用 40 次
- Move as you Say, Interact as you can: Language-Guided Human Motion Generation with Scene AffordanceZan Wang, Yixin Chen, Baoxiong Jia, Puhao Li 等CVPR 2024 · 被引用 38 次
- TACO: Benchmarking Generalizable Bimanual Tool-ACtion-Object UnderstandingYun Liu, Haolin Yang, Xu Si, Ling Liu 等CVPR 2024 · 被引用 12 次
- Harmonizing Stochasticity and Determinism: Scene-responsive Diverse Human Motion PredictionZhenyu Lou, Qiongjie Cui, Tuo Wang, Zhenbo Song 等NeurIPS 2024 · 被引用 10 次
- Decoupled Generative Modeling for Human-Object Interaction SynthesisHwanhee Jung, Seunggwan Lee, Jeongyoon Yoon, SeungHyeon Kim 等CVPR 2026 · 被引用 4 次
它引用的顶会 Paper13
- Learning Trajectory Dependencies for Human Motion PredictionWei Mao, Miaomiao Liu, Mathieu Salzmann, Hongdong LiICCV 2019 · 被引用 534 次
- Resolving 3D Human Pose Ambiguities With 3D Scene ConstraintsMohamed Hassan, Vasileios Choutas, Dimitrios Tzionas, Michael J. BlackICCV 2019 · 被引用 384 次
- Hand-Object Contact Consistency Reasoning for Human Grasps GenerationHanwen Jiang, Shaowei Liu, Jiashun Wang, Xiaolong WangICCV 2021 · 被引用 242 次
- Stochastic Scene-Aware Motion PredictionMohamed Hassan, Duygu Ceylan, Ruben Villegas, Jun Saito 等ICCV 2021 · 被引用 240 次
- Structured Prediction Helps 3D Human Motion ModellingEmre Aksan, Manuel Kaufmann, Otmar HilligesICCV 2019 · 被引用 204 次
相关 Paper
- InterPhys: Physics-aware Human Motion Synthesis in a Dynamic SceneChaoyue Xing, Wei Mao, Miaomiao LiuCVPR 2026 · 被引用 1 次
- Scene-Aware Generative Network for Human Motion SynthesisJingbo Wang, Sijie Yan, Bo Dai, Dahua LinCVPR 2021
- Multimodal Sense-Informed Forecasting of 3D Human MotionsZhenyu Lou, Qiongjie Cui, Haofan Wang, Xu Tang 等CVPR 2024 · 被引用 8 次
- Task-Oriented Human Grasp Synthesis via Context- and Task-Aware DiffusersAn-Lun Liu, Yu-Wei Chao, Yi-Ting ChenICCV 2025 · 被引用 1 次
- Synthesizing Long-Term 3D Human Motion and Interaction in 3D ScenesJiashun Wang, Huazhe Xu, Jingwei Xu, Sifei Liu 等CVPR 2021
