Learning Monocular 3D Reconstruction of Articulated Categories From Motion
Filippos Kokkinos, Iasonas Kokkinos
Abstract
Monocular 3D reconstruction of articulated object categories is challenging due to the lack of training data and the inherent ill-posedness of the problem. In this work we use video self-supervision, forcing the consistency of consecutive 3D reconstructions by a motion-based cycle loss. This largely improves both optimization-based and learningbased 3D mesh reconstruction. We further introduce an interpretable model of 3D template deformations that controls a 3D surface through the displacement of a small number of local, learnable handles. We formulate this operation as a structured layer relying on mesh-laplacian regularization and show that it can be trained in an end-to-end manner. We finally introduce a per-sample numerical optimisation approach that jointly optimises over mesh displacements and cameras within a video, boosting accuracy both for training and also as test time post-processing. While relying exclusively on a small set of videos collected per category for supervision, we obtain state-of-theart reconstructions with diverse shapes, viewpoints and textures for multiple articulated object categories. Supplementary materials, code, and videos are provided on the project page: https://fkokkinos.github.io/ video_3d_reconstruction/ .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 2a90d4d3-a7d3-4812-a370-84bc0a72bc1dCited by top-tier papers12
- CASA: Category-agnostic Skeletal Animal ReconstructionYuefan Wu, Zeyuan Chen, Shaowei Liu, Zhongzheng Ren et al.NeurIPS 2022 · 47 citations
- To The Point: Correspondence-driven monocular 3D category reconstructionFilippos Kokkinos, Iasonas KokkinosNeurIPS 2021 · 26 citations
- LEPARD: Learning Explicit Part Discovery for 3D Articulated Shape ReconstructionDi Liu, Anastasis Stathopoulos, Qilong Zhangli, Yunhe Gao et al.NeurIPS 2023 · 24 citations
- MPI-Flow: Learning Realistic Optical Flow with Multiplane ImagesYingping Liang, Jiaming Liu, Debing Zhang, Ying FuICCV 2023 · 12 citations
- KeyTr: Keypoint Transporter for 3D Reconstruction of Deformable Objects in VideosDavid Novotný, Ignacio Rocco, Samarth Sinha, Alexandre Carlier et al.CVPR 2022 · 11 citations
Builds on10
- Learning to Reconstruct 3D Human Pose and Shape via Model-Fitting in the LoopNikos Kolotouros, Georgios Pavlakos, Michael J. Black, Kostas DaniilidisICCV 2019 · 1,139 citations
- Video Instance SegmentationLinjie Yang, Yuchen Fan, Ning XuICCV 2019 · 615 citations
- Three-D Safari: Learning to Estimate Zebra Pose, Shape, and Texture From Images "In the Wild"Silvia Zuffi, Angjoo Kanazawa, Tanya Y. Berger-Wolf, Michael J. BlackICCV 2019 · 183 citations
- TexturePose: Supervising Human Mesh Estimation With Texture ConsistencyGeorgios Pavlakos, Nikos Kolotouros, Kostas DaniilidisICCV 2019 · 109 citations
- Canonical Surface Mapping via Geometric Cycle ConsistencyNilesh Kulkarni, Shubham Tulsiani, Abhinav GuptaICCV 2019 · 104 citations
Related papers
- Online Adaptation for Consistent Mesh Reconstruction in the WildXueting Li, Sifei Liu, Shalini De Mello, Kihwan Kim et al.NeurIPS 2020 · 62 citations
- SelfRecon: Self Reconstruction Your Digital Avatar from Monocular VideoBoyi Jiang, Yang Hong, Hujun Bao, Juyong ZhangCVPR 2022 · 142 citations
- ViSER: Video-Specific Surface Embeddings for Articulated 3D Shape ReconstructionGengshan Yang, Deqing Sun, Varun Jampani, Daniel Vlasic et al.NeurIPS 2021 · 103 citations
- DensePose 3D: Lifting Canonical Surface Maps of Articulated Objects to the Third DimensionRoman Shapovalov, David Novotný, Benjamin Graham, Patrick Labatut et al.ICCV 2021 · 10 citations
- Weakly-Supervised Mesh-Convolutional Hand Reconstruction in the WildDominik Kulon, Riza Alp Güler, Iasonas Kokkinos, Michael M. Bronstein et al.CVPR 2020
