Perspective from a Broader Context: Can Room Style Knowledge Help Visual Floorplan Localization?
Bolei Chen, Shengsheng Yan, Yongzheng Cui, Jiaxu Kang, Ping Zhong, Jianxin Wang
Abstract
Since a building's floorplan remains consistent over time and is inherently robust to changes in visual appearance, visual Floorplan Localization (FLoc) has received increasing attention from researchers. However, as a compact and minimalist representation of the building's layout, floorplans contain many repetitive structures (e.g., hallways and corners), thus easily result in ambiguous localization. Existing methods either pin their hopes on matching 2D structural cues in floorplans or rely on 3D geometry-constrained visual pre-trainings, ignoring the richer contextual information provided by visual images. In this paper, we suggest using broader visual scene context to empower FLoc algorithms with scene layout priors to eliminate localization uncertainty. In particular, we propose an unsupervised learning technique with clustering constraints to pre-train a room discriminator on self-collected unlabeled room images. Such a discriminator can empirically extract the hidden room type of the observed image and distinguish it from other room types. By injecting the scene context information summarized by the discriminator into an FLoc algorithm, the room style knowledge is effectively exploited to guide definite visual FLoc. We conducted sufficient comparative studies on two standard visual Floc benchmarks. Our experiments show that our approach outperforms state-of-the-art methods and achieves significant improvements in robustness and accuracy.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 098fa769-1a8e-4ba1-a57c-41438cc8f643Cited by top-tier papers1
Ask how each one uses itBuilds on12
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Learning Navigational Visual Representations with Semantic Map SupervisionYicong Hong, Yang Zhou, Ruiyi Zhang, Franck Dernoncourt et al.ICCV 2023 · 56 citations
- Curious Representation Learning for Embodied IntelligenceYilun Du, Chuang Gan, Phillip IsolaICCV 2021 · 50 citations
- LASER: LAtent SpacE Rendering for 2D Visual LocalizationZhixiang Min, Naji Khosravan, Zachary Bessinger, Manjunath Narayana et al.CVPR 2022 · 18 citations
- Supercharging Floorplan Localization with Semantic RaysYuval Grader, Hadar Averbuch-ElorICCV 2025 · 12 citations
Related papers
- Perspective from a Higher Dimension: Can 3D Geometric Priors Help Visual Floorplan Localization?Bolei Chen, Jiaxu Kang, Haonan Yang, Ping Zhong et al.ACM MM 2025
- UnLoc: Leveraging Depth Uncertainties for Floorplan LocalizationMatthias Wüest, Francis Engelmann, Ondrej Miksik, Marc Pollefeys et al.ICLR 2026 · 8 citations
- LaLaLoc: Latent Layout Localisation in Dynamic, Unvisited EnvironmentsHenry Howard-Jenkins, José-Raúl Ruiz-Sarmiento, Victor Adrian PrisacariuICCV 2021 · 31 citations
- SSLayout360: Semi-Supervised Indoor Layout Estimation From 360deg PanoramaPhi Vu TranCVPR 2021
- FloNa: Floor Plan Guided Embodied Visual NavigationJiaxin Li, Weiqi Huang, Zan Wang, Wei Liang et al.AAAI 2025 · 11 citations
