DiffInDScene: Diffusion-Based High-Quality 3D Indoor Scene Generation
Xiaoliang Ju, Zhaoyang Huang, Yijiin Li, Guofeng Zhang, Yu Qiao, Hongsheng Li
Abstract
We present DiffInDScene, a novel framework for tackling the problem of high-quality 3D indoor scene generation, which is challenging due to the complexity and diversity of the indoor scene geometry. Although diffusionbased generative models have previously demonstrated impressive performance in image generation and object-level 3D generation, they have not yet been applied to roomlevel 3D generation due to their computationally intensive costs. In DiffInDScene, we propose a cascaded 3D diffusion pipeline that is efficient and possesses strong generative performance for Truncated Signed Distance Function (TSDF). The whole pipeline is designed to run on a sparse occupancy space in a coarse-to-fine fashion. Inspired by KinectFusion's incremental alignment and fusion of local TSDF volumes, we propose a diffusion-based SDF fusion ⇤ Joint first authorship Please visit our project page for the latest updates: https:// akirahero.github.io/diffindscene/ approach that iteratively diffuses and fuses local TSDF volumes, facilitating the generation of an entire room environment. The generated results demonstrate that our work is capable to achieve high-quality room generation directly in three-dimensional space, starting from scratch. In addition to the scene generation, the final part of DiffInDScene can be used as a post-processing module to refine the 3D reconstruction results from multi-view stereo. According to the user study, the mesh quality generated by our DiffInD-Scene can even outperform the ground truth mesh provided by ScanNet.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e1fe0d9a-82b7-4743-b0ea-36bd6d5ec92bCited by top-tier papers18
- SpaceBlender: Creating Context-Rich Collaborative Spaces Through Generative 3D Scene BlendingNels Numan, Shwetha Rajaram, Balasaravanan Thoravi Kumaravel, Nicolai Marquardt et al.UIST 2024 · 17 citations
- WorldGrow: Generating Infinite 3D WorldSikuang Li, Chen Yang, Jiemin Fang, Taoran Yi et al.AAAI 2026 · 10 citations
- A Global Depth-Range-Free Multi-View Stereo Transformer Network with Pose EmbeddingYitong Dong, Yijin Li, Zhaoyang Huang, Weikang Bian et al.NeurIPS 2024 · 7 citations
- From Programs to Poses: Factored Real-World Scene Generation via Learned Program LibrariesJoy Hsu, Emily Jin, Jiajun Wu, Niloy J. MitraNeurIPS 2025 · 6 citations
- AtlasGS: Atlanta-world Guided Surface Reconstruction with Implicit Structured GaussiansXiyu Zhang, Chong Bao, Yipeng Chen, Hongjia Zhai et al.NeurIPS 2025 · 5 citations
Builds on20
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion ModelsAlexander Quinn Nichol, Prafulla Dhariwal, Aditya Ramesh, Pranav Shyam et al.ICML 2022 · 4,691 citations
- Instant neural graphics primitives with a multiresolution hash encodingThomas Müller, Alex Evans, Christoph Schied, Alexander KellerSIGGRAPH 2022 · 4,089 citations
- Video Diffusion ModelsJonathan Ho, Tim Salimans, Alexey A. Gritsenko, William Chan et al.NeurIPS 2022 · 2,948 citations
Related papers
- DiffuScene: Denoising Diffusion Models for Generative Indoor Scene SynthesisJiapeng Tang, Yinyu Nie, Lev Markhasin, Angela Dai et al.CVPR 2024 · 62 citations
- BlockFusion: Expandable 3D Scene Generation using Latent Tri-plane ExtrapolationZhennan Wu, Yang Li, Han Yan, Taizhang Shang et al.SIGGRAPH 2024 · 28 citations
- PatchScene: Patch-based Voxel Diffusion Model for Large-Scale Scene CompletionQingdong Xu, Jiajun Zhu, Shilin Zhu, Xinjing He et al.CVPR 2026
- Large Scene Generation with Cube-Absorb Discrete DiffusionQianjiang Hu Wei Hu, Wei HuICCV 2025 · 3 citations
- DiffComplete: Diffusion-based Generative 3D Shape CompletionRuihang Chu, Enze Xie, Shentong Mo, Zhenguo Li et al.NeurIPS 2023 · 66 citations
