LASA: Instance Reconstruction from Real Scans using A Large-scale Aligned Shape Annotation Dataset
Haolin Liu, Chongjie Ye, Yinyu Nie, Yingfan He, Xiaoguang Han
Abstract
Instance shape reconstruction from a 3D scene involves recovering the full geometries of multiple objects at the se-mantic instance level. Many methods leverage data-driven learning due to the intricacies of scene complexity and sig-nificant indoor occlusions. Training these methods often requires a large-scale, high-quality dataset with aligned and paired shape annotations with real-world scans. Existing datasets are either synthetic or misaligned, restricting the performance of data-driven methods on real data. To this end, we introduce LASA, a Large-scale Aligned Shape Annotation Dataset comprising 10,412 high-quality CAD annotations aligned with 920 real-world scene scans from ArkitScenes, created manually by professional artists. On this top, we propose a novel Diffusion-based Cross-Modal Shape Reconstruction (DisCo) method. It is empowered by a hybrid feature aggregation design to fuse multi-modal in-puts and recover high-fidelity object geometries (see Fig. 1). Besides, we present an Occupancy-Guided 3D Object De-tection (OccGOD) method and demonstrate that our shape annotations provide scene occupancy clues that can further improve 3D object detection. Supported by LASA, extensive experiments show that our methods achieve state-of-the-art performance in both instance-level scene reconstruction and 3D object detection tasks.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext c2ffcd23-0893-49e6-a3c6-64ec489350d1Cited by top-tier papers2
- Diorama: Unleashing Zero-Shot Single-View 3D Indoor Scene ModelingQirui Wu, Denys Iliash, Daniel Ritchie, Manolis Savva et al.ICCV 2025 · 4 citations
- UniRestore3D: A Scalable Framework For General Shape RestorationYuang Wang, Yujian Zhang, Sida Peng, Xingyi He et al.ICLR 2025
Builds on24
- Elucidating the Design Space of Diffusion-Based Generative ModelsTero Karras, Miika Aittala, Timo Aila, Samuli LaineNeurIPS 2022 · 3,959 citations
- Deep Hough Voting for 3D Object Detection in Point CloudsCharles R. Qi, Or Litany, Kaiming He, Leonidas J. GuibasICCV 2019 · 1,467 citations
- PIFu: Pixel-Aligned Implicit Function for High-Resolution Clothed Human DigitizationShunsuke Saito, Zeng Huang, Ryota Natsume, Shigeo Morishima et al.ICCV 2019 · 1,411 citations
- 3D-FRONT: 3D Furnished Rooms with layOuts and semaNTicsHuan Fu, Bowen Cai, Lin Gao, Lingxiao Zhang et al.ICCV 2021 · 419 citations
- Pix2Vox: Context-Aware 3D Reconstruction From Single and Multi-View ImagesHaozhe Xie, Hongxun Yao, Xiaoshuai Sun, Shangchen Zhou et al.ICCV 2019 · 373 citations
Related papers
- InstaScene: Towards Complete 3D Instance Decomposition and Reconstruction From Cluttered ScenesZesong Yang, Bangbang Yang, Liyuan Cui, Yuewen Ma et al.ICCV 2025 · 2 citations
- UnScene3D: Unsupervised 3D Instance Segmentation for Indoor ScenesDávid Rozenberszki, Or Litany, Angela DaiCVPR 2024 · 25 citations
- DiffComplete: Diffusion-based Generative 3D Shape CompletionRuihang Chu, Enze Xie, Shentong Mo, Zhenguo Li et al.NeurIPS 2023 · 66 citations
- Point-based Instance Completion with Scene ConstraintsWesley Khademi, Fuxin LiICLR 2025
- ASSIST-3D: Adapted Scene Synthesis for Class-Agnostic 3D Instance SegmentationShengchao Zhou, Jiehong Lin, Jiahui Liu, Shizhen Zhao et al.AAAI 2026
