CamNet: Coarse-to-Fine Retrieval for Camera Re-Localization
Mingyu Ding, Zhe Wang, Jiankai Sun, Jianping Shi, Ping Luo
Abstract
Camera re-localization is an important but challenging task in applications like robotics and autonomous driving. Recently, retrieval-based methods have been considered as a promising direction as they can be easily generalized to novel scenes. Despite significant progress has been made, we observe that the performance bottleneck of previous methods actually lies in the retrieval module. These methods use the same features for both retrieval and relative pose regression tasks which have potential conflicts in learning. To this end, here we present a coarse-to-fine retrieval-based deep learning framework, which includes three steps, i.e., image-based coarse retrieval, pose-based fine retrieval and precise relative pose regression. With our carefully designed retrieval module, the relative pose regression task can be surprisingly simpler. We design novel retrieval losses with batch hard sampling criterion and two-stage retrieval to locate samples that adapt to the relative pose regression task. Extensive experiments show that our model (CamNet) outperforms the state-of-the-art methods by a large margin on both indoor and outdoor datasets.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d5917f2c-dccb-4d60-97b4-58181ab9b085Cited by top-tier papers33
- Dual-Resolution Correspondence NetworksXinghui Li, Kai Han, Shuda Li, Victor PrisacariuNeurIPS 2020 · 207 citations
- CroCo: Self-Supervised Pre-training for 3D Vision Tasks by Cross-View CompletionPhilippe Weinzaepfel, Vincent Leroy, Thomas Lucas, Romain Brégier et al.NeurIPS 2022 · 189 citations
- Learning Multi-Scene Absolute Pose Regression with TransformersYoli Shavit, Ron Ferens, Yosi KellerICCV 2021 · 163 citations
- Learning Attribute-driven Disentangled Representations for Interactive Fashion RetrievalYuxin Hou, Eleonora Vig, Michael Donoser, Loris BazzaniICCV 2021 · 58 citations
- Large-Scale Person Detection and Localization using Overhead Fisheye CamerasLu Yang, Liulei Li, Xueshi Xin, Yifan Sun et al.ICCV 2023 · 35 citations
Related papers
- Reloc3r: Large-Scale Training of Relative Camera Pose Regression for Generalizable, Fast, and Accurate Visual LocalizationSiyan Dong, Shuzhe Wang, Shaohui Liu, Lulu Cai et al.CVPR 2025
- From Sparse to Dense: Camera Relocalization with Scene-Specific Detector from Feature Gaussian SplattingZhiwei Huang, Hailin Yu, Yichun Shentu, Jin Yuan et al.CVPR 2025
- KFNet: Learning Temporal Camera Relocalization Using Kalman FilteringLei Zhou, Zixin Luo, Tianwei Shen, Jiahui Zhang et al.CVPR 2020
- Beyond Cross-view Image Retrieval: Highly Accurate Vehicle Localization Using Satellite ImageYujiao Shi, Hongdong LiCVPR 2022 · 81 citations
- RobustLoc: Robust Camera Pose Regression in Challenging Driving EnvironmentsSijie Wang, Qiyu Kang, Rui She, Wee Peng Tay et al.AAAI 2023 · 27 citations
