Structure from Duplicates: Neural Inverse Graphics from a Pile of Objects
Tianhang Cheng, Wei-Chiu Ma, Kaiyu Guan, Antonio Torralba, Shenlong Wang
Abstract
Our world is full of identical objects (.g., cans of coke, cars of same model). These duplicates, when seen together, provide additional and strong cues for us to effectively reason about 3D. Inspired by this observation, we introduce Structure from Duplicates (SfD), a novel inverse graphics framework that reconstructs geometry, material, and illumination from a single image containing multiple identical objects. SfD begins by identifying multiple instances of an object within an image, and then jointly estimates the 6DoF pose for all instances.An inverse graphics pipeline is subsequently employed to jointly reason about the shape, material of the object, and the environment light, while adhering to the shared geometry and material constraint across instances. Our primary contributions involve utilizing object duplicates as a robust prior for single-image inverse graphics and proposing an in-plane rotation-robust Structure from Motion (SfM) formulation for joint 6-DoF object pose estimation. By leveraging multi-view cues from a single image, SfD generates more realistic and detailed 3D reconstructions, significantly outperforming existing single image reconstruction models and multi-view reconstruction approaches with a similar or greater number of observations.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers3
- DreamMat: High-quality PBR Material Generation with Geometry- and Light-aware Diffusion ModelsYuqing Zhang, Yuan Liu, Zhiyu Xie, Lei Yang et al.SIGGRAPH 2024 · 28 citations
- Splat and Replace: 3D Reconstruction with Repetitive ElementsNicolás Violante, Andreas Meuleman, Alban Gauthier, Frédo Durand et al.SIGGRAPH 2025 · 4 citations
- DNF-Intrinsic: Deterministic Noise-Free Diffusion for Indoor Inverse RenderingRongjia Zheng, Qing Zhang, Chengjiang Long, Wei-Shi ZhengICCV 2025 · 2 citations
Builds on26
- Implicit Neural Representations with Periodic Activation FunctionsVincent Sitzmann, Julien N. P. Martel, Alexander W. Bergman, David B. Lindell et al.NeurIPS 2020 · 4,008 citations
- Vision Transformers for Dense PredictionRené Ranftl, Alexey Bochkovskiy, Vladlen KoltunICCV 2021 · 2,647 citations
- NeuS: Learning Neural Implicit Surfaces by Volume Rendering for Multi-view ReconstructionPeng Wang, Lingjie Liu, Yuan Liu, Christian Theobalt et al.NeurIPS 2021 · 2,500 citations
- Multiview Neural Surface Reconstruction by Disentangling Geometry and AppearanceLior Yariv, Yoni Kasten, Dror Moran, Meirav Galun et al.NeurIPS 2020 · 1,010 citations
- Efficient Geometry-aware 3D Generative Adversarial NetworksEric R. Chan, Connor Z. Lin, Matthew A. Chan, Koki Nagano et al.CVPR 2022 · 984 citations
Related papers
- 3DP3: 3D Scene Perception via Probabilistic ProgrammingNishad Gothoskar, Marco F. Cusumano-Towner, Ben Zinberg, Matin Ghavamizadeh et al.NeurIPS 2021 · 59 citations
- Factored-NeuS: Reconstructing Surfaces, Illumination, and Materials of Possibly Glossy ObjectsYue Fan, Ningjing Fan, Ivan Skorokhodov, Oleg Voynov et al.CVPR 2025
- Reconstruct Locally, Localize Globally: A Model Free Method for Object Pose EstimationMing Cai, Ian ReidCVPR 2020
- DiffusionSfM: Predicting Structure and Motion via Ray Origin and Endpoint DiffusionQitao Zhao, Amy Lin, Jeff Tan, Jason Y. Zhang et al.CVPR 2025
- SGS-Intrinsic: Semantic-Invariant Gaussian Splatting for Sparse-View Indoor Inverse RenderingJiahao Niu, Rongjia Zheng, Wenju Xu, Wei-Shi Zheng et al.CVPR 2026 · 1 citation
