Differentiable Blocks World: Qualitative 3D Decomposition by Rendering Primitives
Tom Monnier, Jake Austin, Angjoo Kanazawa, Alexei A. Efros, Mathieu Aubry
摘要
Given a set of calibrated images of a scene, we present an approach that produces a simple, compact, and actionable 3D world representation by means of 3D primitives. While many approaches focus on recovering high-fidelity 3D scenes, we focus on parsing a scene into mid-level 3D representations made of a small set of textured primitives. Such representations are interpretable, easy to manipulate and suited for physics-based simulations. Moreover, unlike existing primitive decomposition methods that rely on 3D input data, our approach operates directly on images through differentiable rendering. Specifically, we model primitives as textured superquadric meshes and optimize their parameters from scratch with an image rendering loss. We highlight the importance of modeling transparency for each primitive, which is critical for optimization and also enables handling varying numbers of primitives. We show that the resulting textured primitives faithfully reconstruct the input images and accurately model the visible 3D points, while providing amodal shape completions of unseen object regions. We compare our approach to the state of the art on diverse scenes from DTU, and demonstrate its robustness on real-life captures from BlendedMVS and Nerfstudio. We also showcase how our results can be used to effortlessly edit a scene or perform physical simulations. Code and video results are available at https://www.tmonnier.com/DBW .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper20
- Disentangled 3D Scene Generation with Layout LearningDave Epstein, Ben Poole, Ben Mildenhall, Alexei A. Efros 等ICML 2024 · 被引用 39 次
- AutoPartGen: Autoregressive 3D Part Generation and DiscoveryMinghao Chen, Jianyuan Wang, Roman Shapovalov, Tom Monnier 等NeurIPS 2025 · 被引用 29 次
- Visual Jenga: Discovering Object Dependencies via Counterfactual InpaintingAnand Bhattad, Konpat Preechakul, Alexei A. EfrosNeurIPS 2025 · 被引用 13 次
- PrimitiveAnything: Human-Crafted 3D Primitive Assembly Generation with Auto-Regressive transformerJingwen Ye, Yuze He, Yanning Zhou, Yiqin Zhu 等SIGGRAPH 2025 · 被引用 5 次
- Residual Primitive Fitting of 3D Shapes with SuperFrustaAditya Ganeshan, Matheus Gadelha, Thibault Groueix, Zhiqin Chen 等CVPR 2026 · 被引用 5 次
它引用的顶会 Paper22
- NeuS: Learning Neural Implicit Surfaces by Volume Rendering for Multi-view ReconstructionPeng Wang, Lingjie Liu, Yuan Liu, Christian Theobalt 等NeurIPS 2021 · 被引用 2,500 次
- Volume Rendering of Neural Implicit SurfacesLior Yariv, Jiatao Gu, Yoni Kasten, Yaron LipmanNeurIPS 2021 · 被引用 1,421 次
- Multiview Neural Surface Reconstruction by Disentangling Geometry and AppearanceLior Yariv, Yoni Kasten, Dror Moran, Meirav Galun 等NeurIPS 2020 · 被引用 1,010 次
- UNISURF: Unifying Neural Implicit Surfaces and Radiance Fields for Multi-View ReconstructionMichael Oechsle, Songyou Peng, Andreas GeigerICCV 2021 · 被引用 885 次
- Nerfstudio: A Modular Framework for Neural Radiance Field DevelopmentMatthew Tancik, Ethan Weber, Evonne Ng, Ruilong Li 等SIGGRAPH 2023 · 被引用 592 次
相关 Paper
- SuperDec: 3D Scene Decomposition with Superquadric PrimitivesElisabetta Fedele, Boyang Sun, Leonidas J. Guibas, Marc Pollefeys 等ICCV 2025 · 被引用 3 次
- Self-Supervised Learning of Hybrid Part-Aware 3D Representations of 2D Gaussians and SuperquadricsZhirui Gao, Renjiao Yi, Yuhang Huang, Wei Chen 等ICCV 2025 · 被引用 2 次
- DualPrim: Compact 3D Reconstruction with Positive and Negative PrimitivesXiaoxu Meng, Zhongmin Chen, Bo Yang, Weikai Chen 等CVPR 2026
- Marching-Primitives: Shape Abstraction from Signed Distance FunctionWeixiao Liu, Yuwei Wu, Sipu Ruan, Gregory S. ChirikjianCVPR 2023
- QUADify: Extracting Meshes with Pixel-Level Details and Materials from ImagesMaximilian Frühauf, Hayko Riemenschneider, Markus Gross, Christopher SchroersCVPR 2024
