Is This the Right Place? Geometric-Semantic Pose Verification for Indoor Visual Localization
Hajime Taira, Ignacio Rocco, Jirí Sedlár, Masatoshi Okutomi, Josef Sivic, Tomás Pajdla, Torsten Sattler, Akihiko Torii
Abstract
Visual localization in large and complex indoor scenes, dominated by weakly textured rooms and repeating geometric patterns, is a challenging problem with high practical relevance for applications such as Augmented Reality and robotics. To handle the ambiguities arising in this scenario, a common strategy is, first, to generate multiple estimates for the camera pose from which a given query image was taken. The pose with the largest geometric consistency with the query image, e.g., in the form of an inlier count, is then selected in a second stage. While a significant amount of research has concentrated on the first stage, there has been considerably less work on the second stage. In this paper, we thus focus on pose verification. We show that combining different modalities, namely appearance, geometry, and semantics, considerably boosts pose verification and consequently pose accuracy. We develop multiple hand-crafted as well as a trainable approach to join into the geometric-semantic verification and show significant improvements over state-of-the-art on a very challenging indoor dataset.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext cf464476-5229-4037-a7be-ffd92ba221b9Cited by top-tier papers8
- Pose Correction for Highly Accurate Visual Localization in Large-scale Indoor SpacesJanghun Hyeon, Joohyung Kim, Nakju Lett DohICCV 2021 · 26 citations
- GLACE: Global Local Accelerated Coordinate EncodingFangjinhua Wang, Xudong Jiang, Silvano Galliani, Christoph Vogel et al.CVPR 2024 · 24 citations
- Scaling Image Geo-Localization to Continent LevelPhilipp Lindenberger, Paul-Edouard Sarlin, Jan Hosang, Marc Pollefeys et al.NeurIPS 2025 · 11 citations
- The Unreasonable Effectiveness of Pre-Trained Features for Camera Pose RefinementGabriele Trivigno, Carlo Masone, Barbara Caputo, Torsten SattlerCVPR 2024 · 10 citations
- Exploring Matching Rates: From Keypoint Selection to Camera RelocalizationHu Lin, Chengjiang Long, Yifeng Fei, Qianchen Xia et al.ACM MM 2024 · 1 citation
Related papers
- Large-Scale Localization Datasets in Crowded Indoor SpacesDonghwan Lee, Soo-Hyun Ryu, Suyong Yeon, Yonghan Lee et al.CVPR 2021
- Local Supports Global: Deep Camera Relocalization With Sequence EnhancementFei Xue, Xin Wang, Zike Yan, Qiuyuan Wang et al.ICCV 2019 · 57 citations
- NormalLoc: Visual Localization on Textureless 3D Models using Surface NormalsJiro Abe, Gaku Nakano, Kazumine OguraICCV 2025 · 2 citations
- CroCoDL: Cross-device Collaborative Dataset for LocalizationHermann Blum, Alessandro Mercurio, Joshua O'Reilly, Tim Engelbracht et al.CVPR 2025
- RGB2LIDAR: Towards Solving Large-Scale Cross-Modal Visual LocalizationNiluthpol Chowdhury Mithun, Karan Sikka, Han-Pang Chiu, Supun Samarasekera et al.ACM MM 2020 · 22 citations
