Connecting the Dots: Floorplan Reconstruction Using Two-Level Queries
Yuanwen Yue, Theodora Kontogianni, Konrad Schindler, Francis Engelmann
Abstract
We address 2D floorplan reconstruction from 3D scans. Existing approaches typically employ heuristically designed multi-stage pipelines. Instead, we formulate floorplan reconstruction as a single-stage structured prediction task: find a variable-size set of polygons, which in turn are variable-length sequences of ordered vertices. To solve it we develop a novel Transformer architecture that generates polygons of multiple rooms in parallel, in a holistic manner without hand-crafted intermediate stages. The model features two-level queries for polygons and corners, and includes polygon matching to make the network end-to-end trainable. Our method achieves a new state-of-the-art for two challenging datasets, Struc-tured3D and SceneCAD, along with significantly faster inference than previous methods. Moreover, it can readily be extended to predict additional information, i.e., semantic room types and architectural elements like doors and windows. Our code and models are available at: https://github.com/ywyue/RoomFormer .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext cddf0c32-dbe9-4d87-a66c-6da6547dada2Cited by top-tier papers18
- OpenMask3D: Open-Vocabulary 3D Instance SegmentationAyça Takmaz, Elisabetta Fedele, Robert W. Sumner, Marc Pollefeys et al.NeurIPS 2023 · 389 citations
- SpatialLM: Training Large Language Models for Structured Indoor ModelingYongsen Mao, Junhao Zhong, Chuan Fang, Jia Zheng et al.NeurIPS 2025 · 89 citations
- OpenNeRF: Open Set 3D Neural Scene Segmentation with Pixel-Wise Features and Rendered Novel ViewsFrancis Engelmann, Fabian Manhardt, Michael Niemeyer, Keisuke Tateno et al.ICLR 2024 · 69 citations
- PolyDiffuse: Polygonal Shape Reconstruction via Guided Set Diffusion ModelsJiacheng Chen, Ruizhi Deng, Yasutaka FurukawaNeurIPS 2023 · 52 citations
- SLIBO-Net: Floorplan Reconstruction via Slicing Box Representation with Local Geometry RegularizationJheng-Wei Su, Kuei-Yu Tung, Chi-Han Peng, Peter Wonka et al.NeurIPS 2023 · 12 citations
Builds on14
- SegFormer: Simple and Efficient Design for Semantic Segmentation with TransformersEnze Xie, Wenhai Wang, Zhiding Yu, Anima Anandkumar et al.NeurIPS 2021 · 9,661 citations
- Deformable DETR: Deformable Transformers for End-to-End Object DetectionXizhou Zhu, Weijie Su, Lewei Lu, Bin Li et al.ICLR 2021 · 7,353 citations
- Per-Pixel Classification is Not All You Need for Semantic SegmentationBowen Cheng, Alexander G. Schwing, Alexander KirillovNeurIPS 2021 · 2,196 citations
- TrackFormer: Multi-Object Tracking with TransformersTim Meinhardt, Alexander Kirillov, Laura Leal-Taixé, Christoph FeichtenhoferCVPR 2022 · 927 citations
- Text Spotting TransformersXiang Zhang, Yongwen Su, Subarna Tripathi, Zhuowen TuCVPR 2022 · 125 citations
Related papers
- FloorPlanFormer: Multi-Task Transformer Network for Floor Plan Recognition with Outer-to-Inner Feature RefinementYun Liang, Zihao Wu, Run Zheng, Shuai Xie et al.AAAI 2026
- CAGE: Continuity-Aware edGE Network Unlocks Robust Floorplan ReconstructionYiyi Liu, Chunyang Liu, Bohan Wang, Weiqin Jiao et al.NeurIPS 2025 · 7 citations
- Raster2Seq: Polygon Sequence Generation for Floorplan ReconstructionHao Phung, Hadar Averbuch-ElorSIGGRAPH 2026 · 2 citations
- HEAT: Holistic Edge Attention Transformer for Structured ReconstructionJiacheng Chen, Yiming Qian, Yasutaka FurukawaCVPR 2022 · 35 citations
- GGPT: Geometry-Grounded Point TransformerYutong Chen, Yiming Wang, Xucong Zhang, Sergey Prokudin et al.CVPR 2026 · 2 citations
