PhysPT: Physics-aware Pretrained Transformer for Estimating Human Dynamics from Monocular Videos
Yufei Zhang, Jeffrey O. Kephart, Zijun Cui, Qiang Ji
摘要
While current methods have shown promising progress on estimating 3D human motion from monocular videos, their motion estimates are often physically unrealistic be-cause they mainly consider kinematics. In this paper, we in-troduce Physics-aware Pretrained Transformer (PhysPT), which improves kinematics-based motion estimates and in-fers motion forces. PhysPT exploits a Transformer encoder-decoder backbone to effectively learn human dynamics in a self-supervised manner. Moreover, it incorporates physics principles governing human motion. Specifically, we build a physics-based body representation and contact force model. We leverage them to impose novel physics-inspired training losses (i.e., force loss, contact loss, and Euler-Lagrange loss), enabling PhysPT to capture physical properties of the human body and the forces it experiences. Experiments demonstrate that, once trained, PhysPT can be directly ap-plied to kinematics-based estimates to significantly enhance their physical plausibility and generate favourable motion forces. Furthermore, we show that these physically meaningful quantities translate into improved accuracy of an important downstream task: human action recognition.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper15
- 🎧MOSPA: Human Motion Generation Driven by Spatial AudioShuyang Xu, Zhiyang Dou, Mingyi Shi, Liang Pan 等NeurIPS 2025 · 被引用 13 次
- Seeing the Wind from a Falling LeafZhiyuan Gao, Jiageng Mao, Hong-Xing Yu, Haozhe Lou 等NeurIPS 2025 · 被引用 10 次
- Optimal-state Dynamics Estimation for Physics-based Human Motion Capture from VideosCuong Le, John Viktor Johansson, Manon Kok, Bastian WandtNeurIPS 2024 · 被引用 8 次
- A Plug-And-Play Physical Motion Restoration Approach for In-The-Wild High-Difficulty MotionsYouliang Zhang, Ronghui Li, Yachao Zhang, Liang Pan 等ICCV 2025 · 被引用 4 次
- PAD-Hand: Physics-Aware Diffusion for Hand Motion RecoveryElkhan Ismayilzada, Yufei Zhang, Zijun CuiCVPR 2026 · 被引用 4 次
它引用的顶会 Paper54
- AMASS: Archive of Motion Capture As Surface ShapesNaureen Mahmood, Nima Ghorbani, Nikolaus F. Troje, Gerard Pons-Moll 等ICCV 2019 · 被引用 1,784 次
- Learning to Reconstruct 3D Human Pose and Shape via Model-Fitting in the LoopNikos Kolotouros, Georgios Pavlakos, Michael J. Black, Kostas DaniilidisICCV 2019 · 被引用 1,139 次
- Channel-wise Topology Refinement Graph Convolution for Skeleton-Based Action RecognitionYuxin Chen, Ziqi Zhang, Chunfeng Yuan, Bing Li 等ICCV 2021 · 被引用 871 次
- PARE: Part Attention Regressor for 3D Human Body EstimationMuhammed Kocabas, Chun-Hao P. Huang, Otmar Hilliges, Michael J. BlackICCV 2021 · 被引用 509 次
- Exploiting Spatial-Temporal Relationships for 3D Pose Estimation via Graph Convolutional NetworksYujun Cai, Liuhao Ge, Jun Liu, Jianfei Cai 等ICCV 2019 · 被引用 504 次
相关 Paper
- Trajectory Optimization for Physics-Based Reconstruction of 3d Human Pose from Monocular VideoErik Gärtner, Mykhaylo Andriluka, Hongyi Xu, Cristian SminchisescuCVPR 2022 · 被引用 31 次
- Physics-based Human Motion Estimation and Synthesis from VideosKevin Xie, Tingwu Wang, Umar Iqbal, Yunrong Guo 等ICCV 2021 · 被引用 102 次
- Learning to Estimate External Forces of Human Motion in VideoNathan Louis, Jason J. Corso, Tylan N. Templin, Travis D. Eliason 等ACM MM 2022 · 被引用 11 次
- 3D Human Pose Estimation with MusclesKevin Zhu, Ali Asghar Mohammadi Nasrabadi, Alexander Wong, John McPheeNeurIPS 2025
- Physics-Augmented Autoencoder for 3D Skeleton-Based Gait RecognitionHongji Guo, Qiang JiICCV 2023 · 被引用 24 次
