DiffInDScene: Diffusion-Based High-Quality 3D Indoor Scene Generation
Xiaoliang Ju, Zhaoyang Huang, Yijiin Li, Guofeng Zhang, Yu Qiao, Hongsheng Li
摘要
We present DiffInDScene, a novel framework for tackling the problem of high-quality 3D indoor scene generation, which is challenging due to the complexity and diversity of the indoor scene geometry. Although diffusionbased generative models have previously demonstrated impressive performance in image generation and object-level 3D generation, they have not yet been applied to roomlevel 3D generation due to their computationally intensive costs. In DiffInDScene, we propose a cascaded 3D diffusion pipeline that is efficient and possesses strong generative performance for Truncated Signed Distance Function (TSDF). The whole pipeline is designed to run on a sparse occupancy space in a coarse-to-fine fashion. Inspired by KinectFusion's incremental alignment and fusion of local TSDF volumes, we propose a diffusion-based SDF fusion ⇤ Joint first authorship Please visit our project page for the latest updates: https:// akirahero.github.io/diffindscene/ approach that iteratively diffuses and fuses local TSDF volumes, facilitating the generation of an entire room environment. The generated results demonstrate that our work is capable to achieve high-quality room generation directly in three-dimensional space, starting from scratch. In addition to the scene generation, the final part of DiffInDScene can be used as a post-processing module to refine the 3D reconstruction results from multi-view stereo. According to the user study, the mesh quality generated by our DiffInD-Scene can even outperform the ground truth mesh provided by ScanNet.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper18
- SpaceBlender: Creating Context-Rich Collaborative Spaces Through Generative 3D Scene BlendingNels Numan, Shwetha Rajaram, Balasaravanan Thoravi Kumaravel, Nicolai Marquardt 等UIST 2024 · 被引用 17 次
- WorldGrow: Generating Infinite 3D WorldSikuang Li, Chen Yang, Jiemin Fang, Taoran Yi 等AAAI 2026 · 被引用 10 次
- A Global Depth-Range-Free Multi-View Stereo Transformer Network with Pose EmbeddingYitong Dong, Yijin Li, Zhaoyang Huang, Weikang Bian 等NeurIPS 2024 · 被引用 7 次
- From Programs to Poses: Factored Real-World Scene Generation via Learned Program LibrariesJoy Hsu, Emily Jin, Jiajun Wu, Niloy J. MitraNeurIPS 2025 · 被引用 6 次
- AtlasGS: Atlanta-world Guided Surface Reconstruction with Implicit Structured GaussiansXiyu Zhang, Chong Bao, Yipeng Chen, Hongjia Zhai 等NeurIPS 2025 · 被引用 5 次
它引用的顶会 Paper20
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion ModelsAlexander Quinn Nichol, Prafulla Dhariwal, Aditya Ramesh, Pranav Shyam 等ICML 2022 · 被引用 4,691 次
- Instant neural graphics primitives with a multiresolution hash encodingThomas Müller, Alex Evans, Christoph Schied, Alexander KellerSIGGRAPH 2022 · 被引用 4,089 次
- Video Diffusion ModelsJonathan Ho, Tim Salimans, Alexey A. Gritsenko, William Chan 等NeurIPS 2022 · 被引用 2,948 次
相关 Paper
- DiffuScene: Denoising Diffusion Models for Generative Indoor Scene SynthesisJiapeng Tang, Yinyu Nie, Lev Markhasin, Angela Dai 等CVPR 2024 · 被引用 62 次
- BlockFusion: Expandable 3D Scene Generation using Latent Tri-plane ExtrapolationZhennan Wu, Yang Li, Han Yan, Taizhang Shang 等SIGGRAPH 2024 · 被引用 28 次
- PatchScene: Patch-based Voxel Diffusion Model for Large-Scale Scene CompletionQingdong Xu, Jiajun Zhu, Shilin Zhu, Xinjing He 等CVPR 2026
- Large Scene Generation with Cube-Absorb Discrete DiffusionQianjiang Hu Wei Hu, Wei HuICCV 2025 · 被引用 3 次
- DiffComplete: Diffusion-based Generative 3D Shape CompletionRuihang Chu, Enze Xie, Shentong Mo, Zhenguo Li 等NeurIPS 2023 · 被引用 66 次
