PPR: Physically Plausible Reconstruction from Monocular Videos
Gengshan Yang, Shuo Yang, John Z. Zhang, Zachary Manchester, Deva Ramanan
摘要
Given casually-captured monocular videos (left), PPR builds 3D models of articulated objects and the surrounding environment. Naive kinematic reconstruction (middle) generates a family of solutions, some containing inconsistent physical support and contact dynamics (blue and green color), such as floating or walking with sliding feet. We show that differentiable physics simulation acts as effective regularizer for improving the physical plausibility of visual reconstruction algorithms. As PPR reconstructs the dynamics scene, it also drives a ragdoll in a physics simulator to track the kinematic reconstruction. This ensures the reconstructions are statically stable with ground contact (right), and the center of mass is projected within the support polygon (marked with red). PPR also reports physics estimations, such as ground reaction forces (red arrows) and center of mass (green arrow).
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper18
- PhysPT: Physics-aware Pretrained Transformer for Estimating Human Dynamics from Monocular VideosYufei Zhang, Jeffrey O. Kephart, Zijun Cui, Qiang JiCVPR 2024 · 被引用 14 次
- Seeing the Wind from a Falling LeafZhiyuan Gao, Jiageng Mao, Hong-Xing Yu, Haozhe Lou 等NeurIPS 2025 · 被引用 10 次
- VAREN: Very Accurate and Realistic Equine NetworkSilvia Zuffi, Ylva Mellbin, Ci Li, Markus Höschle 等CVPR 2024 · 被引用 9 次
- HoliGS: Holistic Gaussian Splatting for Embodied View SynthesisXiaoyuan Wang, Yizhou Zhao, Botao Ye, Xiaojun Shan 等NeurIPS 2025 · 被引用 8 次
- REACTO: Reconstructing Articulated Objects from a Single VideoChaoyue Song, Jiacheng Wei, Chuan Sheng Foo, Guosheng Lin 等CVPR 2024 · 被引用 7 次
它引用的顶会 Paper35
- Vision Transformers for Dense PredictionRené Ranftl, Alexey Bochkovskiy, Vladlen KoltunICCV 2021 · 被引用 2,647 次
- Nerfies: Deformable Neural Radiance FieldsKeunhong Park, Utkarsh Sinha, Jonathan T. Barron, Sofien Bouaziz 等ICCV 2021 · 被引用 1,442 次
- Volume Rendering of Neural Implicit SurfacesLior Yariv, Jiatao Gu, Yoni Kasten, Yaron LipmanNeurIPS 2021 · 被引用 1,421 次
- DROID-SLAM: Deep Visual SLAM for Monocular, Stereo, and RGB-D CamerasZachary Teed, Jia DengNeurIPS 2021 · 被引用 1,248 次
- Multiview Neural Surface Reconstruction by Disentangling Geometry and AppearanceLior Yariv, Yoni Kasten, Dror Moran, Meirav Galun 等NeurIPS 2020 · 被引用 1,010 次
相关 Paper
- Differentiable Dynamics for Articulated 3d Human Motion ReconstructionErik Gärtner, Mykhaylo Andriluka, Erwin Coumans, Cristian SminchisescuCVPR 2022 · 被引用 33 次
- Recovering Physically Plausible Human-Object Interactions from Monocular VideosDingbang Huang, Etienne Vouga, Qixing Huang, Georgios PavlakosCVPR 2026
- SAFT: Shape and Appearance of Fabrics from Template via Differentiable Physical Simulations from Monocular VideoDavid Stotko, Reinhard KleinICCV 2025 · 被引用 1 次
- Trajectory Optimization for Physics-Based Reconstruction of 3d Human Pose from Monocular VideoErik Gärtner, Mykhaylo Andriluka, Hongyi Xu, Cristian SminchisescuCVPR 2022 · 被引用 31 次
- gradSim: Differentiable simulation for system identification and visuomotor controlJ. Krishna Murthy, Miles Macklin, Florian Golemo, Vikram Voleti 等ICLR 2021 · 被引用 130 次
