LASSIE: Learning Articulated Shapes from Sparse Image Ensemble via 3D Part Discovery
Chun-Han Yao, Wei-Chih Hung, Yuanzhen Li, Michael Rubinstein, Ming-Hsuan Yang, Varun Jampani
摘要
Creating high-quality articulated 3D models of animals is challenging either via manual creation or using 3D scanning tools. Therefore, techniques to reconstruct articulated 3D objects from 2D images are crucial and highly useful. In this work, we propose a practical problem setting to estimate 3D pose and shape of animals given only a few (10-30) in-the-wild images of a particular animal species (say, horse). Contrary to existing works that rely on pre-defined template shapes, we do not assume any form of 2D or 3D ground-truth annotations, nor do we leverage any multi-view or temporal information. Moreover, each input image ensemble can contain animal instances with varying poses, backgrounds, illuminations, and textures. Our key insight is that 3D parts have much simpler shape compared to the overall animal and that they are robust w.r.t. animal pose articulations. Following these insights, we propose LASSIE, a novel optimization framework which discovers 3D parts in a self-supervised manner with minimal user intervention. A key driving force behind LASSIE is the enforcing of 2D-3D part consistency using self-supervisory deep features. Experiments on Pascal-Part and self-collected in-the-wild animal datasets demonstrate considerably better 3D reconstructions as well as both 2D and 3D part discovery compared to prior arts. Project page: chhankyao.github.io/lassie/
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper27
- LEPARD: Learning Explicit Part Discovery for 3D Articulated Shape ReconstructionDi Liu, Anastasis Stathopoulos, Qilong Zhangli, Yunhe Gao 等NeurIPS 2023 · 被引用 24 次
- Template-free Articulated Gaussian Splatting for Real-time Reposable Dynamic View SynthesisDiwen Wan, Yuxiang Wang, Ruijie Lu, Gang ZengNeurIPS 2024 · 被引用 18 次
- S3O: A Dual-Phase Approach for Reconstructing Dynamic Shape and Skeleton of Articulated Objects from Single Monocular VideoHao Zhang, Fang Li, Samyak Rawlekar, Narendra AhujaICML 2024 · 被引用 13 次
- MoCapAnything: Unified 3D Motion Capture for Arbitrary Skeletons from Monocular VideosKehong Gong, Zhengyu Wen, Xiaoyu He, Mingxi Xu 等CVPR 2026 · 被引用 8 次
- Primitive-Based 3D Human-Object Interaction Modelling and ProgrammingSiqi Liu, Yong-Lu Li, Zhou Fang, Xinpeng Liu 等AAAI 2024 · 被引用 8 次
它引用的顶会 Paper20
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou 等ICCV 2021 · 被引用 8,921 次
- Learning to Reconstruct 3D Human Pose and Shape via Model-Fitting in the LoopNikos Kolotouros, Georgios Pavlakos, Michael J. Black, Kostas DaniilidisICCV 2019 · 被引用 1,139 次
- Soft Rasterizer: A Differentiable Renderer for Image-Based 3D ReasoningShichen Liu, Weikai Chen, Tianye Li, Hao LiICCV 2019 · 被引用 789 次
- NeRD: Neural Reflectance Decomposition from Image CollectionsMark Boss, Raphael Braun, Varun Jampani, Jonathan T. Barron 等ICCV 2021 · 被引用 608 次
相关 Paper
- Hi-LASSIE: High-Fidelity Articulated Shape and Skeleton Discovery from Sparse Image EnsembleChun-Han Yao, Wei-Chih Hung, Yuanzhen Li, Michael Rubinstein 等CVPR 2023
- Learning Articulated Shape with Keypoint Pseudo-Labels from Web ImagesAnastasis Stathopoulos, Georgios Pavlakos, Ligong Han, Dimitris N. MetaxasCVPR 2023
- ARTIC3D: Learning Robust Articulated 3D Shapes from Noisy Web Image CollectionsChun-Han Yao, Amit Raj, Wei-Chih Hung, Michael Rubinstein 等NeurIPS 2023 · 被引用 25 次
- Discovering 3D Parts from Image CollectionsChun-Han Yao, Wei-Chih Hung, Varun Jampani, Ming-Hsuan YangICCV 2021 · 被引用 21 次
- Online Adaptation for Consistent Mesh Reconstruction in the WildXueting Li, Sifei Liu, Shalini De Mello, Kihwan Kim 等NeurIPS 2020 · 被引用 62 次
