Differentiable Blocks World: Qualitative 3D Decomposition by Rendering Primitives
Tom Monnier, Jake Austin, Angjoo Kanazawa, Alexei A. Efros, Mathieu Aubry
Abstract
Given a set of calibrated images of a scene, we present an approach that produces a simple, compact, and actionable 3D world representation by means of 3D primitives. While many approaches focus on recovering high-fidelity 3D scenes, we focus on parsing a scene into mid-level 3D representations made of a small set of textured primitives. Such representations are interpretable, easy to manipulate and suited for physics-based simulations. Moreover, unlike existing primitive decomposition methods that rely on 3D input data, our approach operates directly on images through differentiable rendering. Specifically, we model primitives as textured superquadric meshes and optimize their parameters from scratch with an image rendering loss. We highlight the importance of modeling transparency for each primitive, which is critical for optimization and also enables handling varying numbers of primitives. We show that the resulting textured primitives faithfully reconstruct the input images and accurately model the visible 3D points, while providing amodal shape completions of unseen object regions. We compare our approach to the state of the art on diverse scenes from DTU, and demonstrate its robustness on real-life captures from BlendedMVS and Nerfstudio. We also showcase how our results can be used to effortlessly edit a scene or perform physical simulations. Code and video results are available at https://www.tmonnier.com/DBW .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers20
- Disentangled 3D Scene Generation with Layout LearningDave Epstein, Ben Poole, Ben Mildenhall, Alexei A. Efros et al.ICML 2024 · 39 citations
- AutoPartGen: Autoregressive 3D Part Generation and DiscoveryMinghao Chen, Jianyuan Wang, Roman Shapovalov, Tom Monnier et al.NeurIPS 2025 · 29 citations
- Visual Jenga: Discovering Object Dependencies via Counterfactual InpaintingAnand Bhattad, Konpat Preechakul, Alexei A. EfrosNeurIPS 2025 · 13 citations
- PrimitiveAnything: Human-Crafted 3D Primitive Assembly Generation with Auto-Regressive transformerJingwen Ye, Yuze He, Yanning Zhou, Yiqin Zhu et al.SIGGRAPH 2025 · 5 citations
- Residual Primitive Fitting of 3D Shapes with SuperFrustaAditya Ganeshan, Matheus Gadelha, Thibault Groueix, Zhiqin Chen et al.CVPR 2026 · 5 citations
Builds on22
- NeuS: Learning Neural Implicit Surfaces by Volume Rendering for Multi-view ReconstructionPeng Wang, Lingjie Liu, Yuan Liu, Christian Theobalt et al.NeurIPS 2021 · 2,500 citations
- Volume Rendering of Neural Implicit SurfacesLior Yariv, Jiatao Gu, Yoni Kasten, Yaron LipmanNeurIPS 2021 · 1,421 citations
- Multiview Neural Surface Reconstruction by Disentangling Geometry and AppearanceLior Yariv, Yoni Kasten, Dror Moran, Meirav Galun et al.NeurIPS 2020 · 1,010 citations
- UNISURF: Unifying Neural Implicit Surfaces and Radiance Fields for Multi-View ReconstructionMichael Oechsle, Songyou Peng, Andreas GeigerICCV 2021 · 885 citations
- Nerfstudio: A Modular Framework for Neural Radiance Field DevelopmentMatthew Tancik, Ethan Weber, Evonne Ng, Ruilong Li et al.SIGGRAPH 2023 · 592 citations
Related papers
- SuperDec: 3D Scene Decomposition with Superquadric PrimitivesElisabetta Fedele, Boyang Sun, Leonidas J. Guibas, Marc Pollefeys et al.ICCV 2025 · 3 citations
- Self-Supervised Learning of Hybrid Part-Aware 3D Representations of 2D Gaussians and SuperquadricsZhirui Gao, Renjiao Yi, Yuhang Huang, Wei Chen et al.ICCV 2025 · 2 citations
- DualPrim: Compact 3D Reconstruction with Positive and Negative PrimitivesXiaoxu Meng, Zhongmin Chen, Bo Yang, Weikai Chen et al.CVPR 2026
- Marching-Primitives: Shape Abstraction from Signed Distance FunctionWeixiao Liu, Yuwei Wu, Sipu Ruan, Gregory S. ChirikjianCVPR 2023
- QUADify: Extracting Meshes with Pixel-Level Details and Materials from ImagesMaximilian Frühauf, Hayko Riemenschneider, Markus Gross, Christopher SchroersCVPR 2024
