Simple but Effective Triplet-Based Compression Strategies for Compact Visual Localization
Torsten Sattler, Zuzana Kukelova
Abstract
Visual localization, i.e., the problem of estimating the camera pose from which an image was taken, is an important part of applications such as augmented reality and autonomous robots. Many of these applications require a compact memory footprint. Thus, a considerable amount of work has been spent on designing memory-efficient scene representations for visual localization. In this paper, we focus on compressing the 3D structure of the scene by selecting a subset of points from a Structure-from-Motion (SfM) point cloud. In contrast to prior work, which aims to solve (complex) optimization problems, we propose a simple strategy that is almost trivial to implement. Our compression strategy is based on the idea of selecting triplets of points such that the camera pose of each database image (used to build the SfM point cloud) can be accurately estimated from these triplets. Despite its simplicity, our strategy performs similarly to or better than current state-of-theart structure compression approaches. Combined with standard product quantization approaches to compress feature descriptors, our approach compares favorably with recent learning-based approaches for compact visual localization.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 1dde51e3-905d-41cb-aeb4-339ee6d9310aBuilds on27
- LightGlue: Local Feature Matching at Light SpeedPhilipp Lindenberger, Paul-Edouard Sarlin, Marc PollefeysICCV 2023 · 936 citations
- DISK: Learning local features with policy gradientMichal J. Tyszkiewicz, Pascal Fua, Eduard TrullsNeurIPS 2020 · 652 citations
- Rethinking Visual Geo-localization for Large-Scale ApplicationsGabriele Moreno Berton, Carlo Masone, Barbara CaputoCVPR 2022 · 235 citations
- AtLoc: Attention Guided Camera LocalizationBing Wang, Changhao Chen, Chris Xiaoxuan Lu, Peijun Zhao et al.AAAI 2020 · 189 citations
- Learning Multi-Scene Absolute Pose Regression with TransformersYoli Shavit, Ron Ferens, Yosi KellerICCV 2021 · 163 citations
Related papers
- SceneSqueezer: Learning to Compress Scene for Camera RelocalizationLuwei Yang, Rakesh Shrestha, Wenbo Li, Shuaicheng Liu et al.CVPR 2022 · 29 citations
- How Privacy-Preserving Are Line Clouds? Recovering Scene Details From 3D LinesKunal Chelani, Fredrik Kahl, Torsten SattlerCVPR 2021
- SplatLoc: 3D Gaussian Splatting-based Visual Localization for Augmented RealityHongjia Zhai, Xiyu Zhang, Boming Zhao, Hai Li et al.IEEE VR 2025 · 29 citations
- Paired-Point Lifting for Enhanced Privacy-Preserving Visual LocalizationChunghwan Lee, Jaihoon Kim, Chanhyuk Yun, Je Hyeong HongCVPR 2023
- Cascaded Parallel Filtering for Memory-Efficient Image-Based LocalizationWentao Cheng, Weisi Lin, Kan Chen, Xinfeng ZhangICCV 2019 · 28 citations
