Root Pose Decomposition Towards Generic Non-rigid 3D Reconstruction with Monocular Videos
Yikai Wang, Yinpeng Dong, Fuchun Sun, Xiao Yang
摘要
This work focuses on the 3D reconstruction of non-rigid objects based on monocular RGB video sequences. Concretely, we aim at building high-fidelity models for generic object categories and casually captured scenes. To this end, we do not assume known root poses of objects, and do not utilize category-specific templates or dense pose priors. The key idea of our method, Root Pose Decomposition (RPD), is to maintain a per-frame root pose transformation, meanwhile building a dense field with local transformations to rectify the root pose. The optimization of local transformations is performed by point registration to the canonical space. We also adapt RPD to multi-object scenarios with object occlusions and individual differences. As a result, RPD allows non-rigid 3D reconstruction for complicated scenarios containing objects with large deformations, complex motion patterns, occlusions, and scale diversities of different individuals. Such a pipeline potentially scales to diverse sets of objects in the wild. We experimentally show that RPD surpasses state-of-the-art methods on the challenging DAVIS, OVIS, and AMA datasets. We provide video results in https://rpd-share.github.io .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Vidu4D: Single Generated Video to High-Fidelity 4D Reconstruction with Dynamic Gaussian SurfelsYikai Wang, Xinzhou Wang, Zilong Chen, Zhengyi Wang 等NeurIPS 2024 · 被引用 40 次
- SV-GS: Sparse View 4D Reconstruction with Skeleton-Driven Gaussian SplattingJun-Jee Chao, Volkan IslerCVPR 2026 · 被引用 1 次
- OSN: Infinite Representations of Dynamic 3D Scenes from Monocular VideosZiyang Song, Jinxi Li, Bo YangICML 2024
它引用的顶会 Paper14
- NeuS: Learning Neural Implicit Surfaces by Volume Rendering for Multi-view ReconstructionPeng Wang, Lingjie Liu, Yuan Liu, Christian Theobalt 等NeurIPS 2021 · 被引用 2,500 次
- Nerfies: Deformable Neural Radiance FieldsKeunhong Park, Utkarsh Sinha, Jonathan T. Barron, Sofien Bouaziz 等ICCV 2021 · 被引用 1,442 次
- Multiview Neural Surface Reconstruction by Disentangling Geometry and AppearanceLior Yariv, Yoni Kasten, Dror Moran, Meirav Galun 等NeurIPS 2020 · 被引用 1,010 次
- Lepard: Learning partial point cloud matching in rigid and deformable scenesYang Li, Tatsuya HaradaCVPR 2022 · 被引用 163 次
- Continuous Surface EmbeddingsNatalia Neverova, David Novotný, Marc Szafraniec, Vasil Khalidov 等NeurIPS 2020 · 被引用 116 次
相关 Paper
- Total-Recon: Deformable Scene Reconstruction for Embodied View SynthesisChonghyuk Song, Gengshan Yang, Kangle Deng, Jun-Yan Zhu 等ICCV 2023 · 被引用 27 次
- Towards Robust and Smooth 3D Multi-Person Pose Estimation from Monocular Videos in the WildSungchan Park, Eunyi You, Inhoe Lee, Joonseok LeeICCV 2023 · 被引用 16 次
- Free-Moving Object Reconstruction and Pose Estimation with Virtual CameraHaixin Shi, Yinlin Hu, Daniel Koguciuk, Juan-Ting Lin 等AAAI 2025 · 被引用 2 次
- Class-agnostic Reconstruction of Dynamic Objects from VideosZhongzheng Ren, Xiaoming Zhao, Alexander G. SchwingNeurIPS 2021 · 被引用 11 次
- DreamScene4D: Dynamic Multi-Object Scene Generation from Monocular VideosWen-Hsuan Chu, Lei Ke, Katerina FragkiadakiNeurIPS 2024 · 被引用 75 次
