LASR: Learning Articulated Shape Reconstruction From a Monocular Video
Gengshan Yang, Deqing Sun, Varun Jampani, Daniel Vlasic, Forrester Cole, Huiwen Chang, Deva Ramanan, William T. Freeman, Ce Liu
Abstract
Google Research t=4 t=8 t=12 t=16 LASR (Ours) PIFuHD SMPLify-X VIBE LASR (Ours) UMR-horse 0 °60 °A-CSM (camel template) SMALify horse t=20 t=40 t=60 t=80 Figure 1 . Top: Sample input video frames and articulated shapes recovered by our method (LASR). Bottom: Comparison with existing methods, where the input to each method (either video or image) is denoted at the top left, and the shape template being used is denoted at the bottom right of each result. Many existing approaches on nonrigid shape reconstruction heavily rely on category-specific 3D shape templates, such as SMPL for human [33, 35] and SMAL for quadrupeds [6, 58] . In contrast, LASR jointly recovers the object shape, articulation, and camera parameters from a monocular video without using category-specific shape templates. By relying on generic shape and motion priors, LASR applies to a wider range of nonrigid shapes and yields high-fidelity 3D reconstructions: It recovers both humps of the camel, which are missing from other methods. It also recovers the silk ribbon of the dancer (as denoted by the blue box), which confuses SMPLify-X and VIBE as the right arm. Please refer to video results on our project page.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 8a9ccdc9-870c-4187-8bb9-104e0f38fa60Cited by top-tier papers63
- Tracking Everything Everywhere All at OnceQianqian Wang, Yen-Yu Chang, Ruojin Cai, Zhengqi Li et al.ICCV 2023 · 238 citations
- Monocular Dynamic View Synthesis: A Reality CheckHang Gao, Ruilong Li, Shubham Tulsiani, Bryan Russell et al.NeurIPS 2022 · 235 citations
- Kubric: A scalable dataset generatorKlaus Greff, Francois Belletti, Lucas Beyer, Carl Doersch et al.CVPR 2022 · 183 citations
- NeRS: Neural Reflectance Surfaces for Sparse-view 3D Reconstruction in the WildJason Y. Zhang, Gengshan Yang, Shubham Tulsiani, Deva RamananNeurIPS 2021 · 180 citations
- L4GM: Large 4D Gaussian Reconstruction ModelJiawei Ren, Cheng Xie, Ashkan Mirzaei, Hanxue Liang et al.NeurIPS 2024 · 173 citations
Builds on12
- PIFu: Pixel-Aligned Implicit Function for High-Resolution Clothed Human DigitizationShunsuke Saito, Zeng Huang, Ryota Natsume, Shigeo Morishima et al.ICCV 2019 · 1,411 citations
- Soft Rasterizer: A Differentiable Renderer for Image-Based 3D ReasoningShichen Liu, Weikai Chen, Tianye Li, Hao LiICCV 2019 · 789 citations
- Point2Mesh: a self-prior for deformable meshesRana Hanocka, Gal Metzer, Raja Giryes, Daniel Cohen-OrSIGGRAPH 2020 · 243 citations
- Motion-Attentive Transition for Zero-Shot Video Object SegmentationTianfei Zhou, Shunzhou Wang, Yi Zhou, Yazhou Yao et al.AAAI 2020 · 210 citations
- Three-D Safari: Learning to Estimate Zebra Pose, Shape, and Texture From Images "In the Wild"Silvia Zuffi, Angjoo Kanazawa, Tanya Y. Berger-Wolf, Michael J. BlackICCV 2019 · 183 citations
Related papers
- GART: Gaussian Articulated Template ModelsJiahui Lei, Yufu Wang, Georgios Pavlakos, Lingjie Liu et al.CVPR 2024
- Hi-LASSIE: High-Fidelity Articulated Shape and Skeleton Discovery from Sparse Image EnsembleChun-Han Yao, Wei-Chih Hung, Yuanzhen Li, Michael Rubinstein et al.CVPR 2023
- ViSER: Video-Specific Surface Embeddings for Articulated 3D Shape ReconstructionGengshan Yang, Deqing Sun, Varun Jampani, Daniel Vlasic et al.NeurIPS 2021 · 103 citations
- MultiGO: Towards Multi-level Geometry Learning for Monocular 3D Textured Human ReconstructionGangjian Zhang, Nanjie Yao, Shunsi Zhang, Hanfeng Zhao et al.CVPR 2025
- ShapeClipper: Scalable 3D Shape Learning from Single-View Images via Geometric and CLIP-Based ConsistencyZixuan Huang, Varun Jampani, Anh Thai, Yuanzhen Li et al.CVPR 2023
