From Image Collections to Point Clouds With Self-Supervised Shape and Pose Networks
Navaneet K. L., Ansu Mathew, Shashank Kashyap, Wei-Chih Hung, Varun Jampani, R. Venkatesh Babu
Abstract
Reconstructing 3D models from 2D images is one of the fundamental problems in computer vision. In this work, we propose a deep learning technique for 3D object reconstruction from a single image. Contrary to recent works that either use 3D supervision or multi-view supervision, we use only single view images with no pose information during training as well. This makes our approach more practical requiring only an image collection of an object category and the corresponding silhouettes. We learn both 3D point cloud reconstruction and pose estimation networks in a selfsupervised manner, making use of differentiable point cloud renderer to train with 2D supervision. A key novelty of the proposed technique is to impose 3D geometric reasoning into predicted 3D point clouds by rotating them with randomly sampled poses and then enforcing cycle consistency on both 3D reconstructions and poses. In addition, using single-view supervision allows us to do test-time optimization on a given test image. Experiments on the synthetic ShapeNet and real-world Pix3D datasets demonstrate that our approach, despite using less supervision, can achieve competitive performance compared to pose-supervised and multi-view supervised approaches.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 42100852-ac20-47cf-8d61-a76d73b74108Cited by top-tier papers12
- PointGPT: Auto-regressively Generative Pre-training from Point CloudsGuangyan Chen, Meiling Wang, Yi Yang, Kai Yu et al.NeurIPS 2023 · 219 citations
- MilliPCD: Beyond Traditional Vision Indoor Point Cloud Generation via Handheld Millimeter-Wave DevicesPingping Cai, Sanjib SurUbiComp 2023 · 15 citations
- Learning Canonical 3D Object Representation for Fine-Grained RecognitionSunghun Joung, Seungryong Kim, Minsu Kim, Ig-Jae Kim et al.ICCV 2021 · 14 citations
- Single View Point Cloud Generation via Unified 3D PrototypeYu Lin, Yigong Wang, Yi-Fan Li, Zhuoyi Wang et al.AAAI 2021 · 11 citations
- Self-Supervised Object Detection via Generative Image SynthesisSiva Karthik Mustikovela, Shalini De Mello, Aayush Prakash, Umar Iqbal et al.ICCV 2021 · 3 citations
Related papers
- Single Image Shape-from-SilhouettesYawen Lu, Yuxing Wang, Guoyu LuACM MM 2020 · 6 citations
- ShapeClipper: Scalable 3D Shape Learning from Single-View Images via Geometric and CLIP-Based ConsistencyZixuan Huang, Varun Jampani, Anh Thai, Yuanzhen Li et al.CVPR 2023
- Discovering 3D Parts from Image CollectionsChun-Han Yao, Wei-Chih Hung, Varun Jampani, Ming-Hsuan YangICCV 2021 · 21 citations
- SDF-SRN: Learning Signed Distance 3D Object Reconstruction from Static ImagesChen-Hsuan Lin, Chaoyang Wang, Simon LuceyNeurIPS 2020 · 125 citations
- Leveraging SE(3) Equivariance for Self-supervised Category-Level Object Pose Estimation from Point CloudsXiaolong Li, Yijia Weng, Li Yi, Leonidas J. Guibas et al.NeurIPS 2021 · 61 citations
