Perspective from a Broader Context: Can Room Style Knowledge Help Visual Floorplan Localization?
Bolei Chen, Shengsheng Yan, Yongzheng Cui, Jiaxu Kang, Ping Zhong, Jianxin Wang
摘要
Since a building's floorplan remains consistent over time and is inherently robust to changes in visual appearance, visual Floorplan Localization (FLoc) has received increasing attention from researchers. However, as a compact and minimalist representation of the building's layout, floorplans contain many repetitive structures (e.g., hallways and corners), thus easily result in ambiguous localization. Existing methods either pin their hopes on matching 2D structural cues in floorplans or rely on 3D geometry-constrained visual pre-trainings, ignoring the richer contextual information provided by visual images. In this paper, we suggest using broader visual scene context to empower FLoc algorithms with scene layout priors to eliminate localization uncertainty. In particular, we propose an unsupervised learning technique with clustering constraints to pre-train a room discriminator on self-collected unlabeled room images. Such a discriminator can empirically extract the hidden room type of the observed image and distinguish it from other room types. By injecting the scene context information summarized by the discriminator into an FLoc algorithm, the room style knowledge is effectively exploited to guide definite visual FLoc. We conducted sufficient comparative studies on two standard visual Floc benchmarks. Our experiments show that our approach outperforms state-of-the-art methods and achieves significant improvements in robustness and accuracy.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper12
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- Learning Navigational Visual Representations with Semantic Map SupervisionYicong Hong, Yang Zhou, Ruiyi Zhang, Franck Dernoncourt 等ICCV 2023 · 被引用 56 次
- Curious Representation Learning for Embodied IntelligenceYilun Du, Chuang Gan, Phillip IsolaICCV 2021 · 被引用 50 次
- LASER: LAtent SpacE Rendering for 2D Visual LocalizationZhixiang Min, Naji Khosravan, Zachary Bessinger, Manjunath Narayana 等CVPR 2022 · 被引用 18 次
- Supercharging Floorplan Localization with Semantic RaysYuval Grader, Hadar Averbuch-ElorICCV 2025 · 被引用 12 次
相关 Paper
- Perspective from a Higher Dimension: Can 3D Geometric Priors Help Visual Floorplan Localization?Bolei Chen, Jiaxu Kang, Haonan Yang, Ping Zhong 等ACM MM 2025
- UnLoc: Leveraging Depth Uncertainties for Floorplan LocalizationMatthias Wüest, Francis Engelmann, Ondrej Miksik, Marc Pollefeys 等ICLR 2026 · 被引用 8 次
- LaLaLoc: Latent Layout Localisation in Dynamic, Unvisited EnvironmentsHenry Howard-Jenkins, José-Raúl Ruiz-Sarmiento, Victor Adrian PrisacariuICCV 2021 · 被引用 31 次
- SSLayout360: Semi-Supervised Indoor Layout Estimation From 360deg PanoramaPhi Vu TranCVPR 2021
- FloNa: Floor Plan Guided Embodied Visual NavigationJiaxin Li, Weiqi Huang, Zan Wang, Wei Liang 等AAAI 2025 · 被引用 11 次
