Beyond Cross-view Image Retrieval: Highly Accurate Vehicle Localization Using Satellite Image
Yujiao Shi, Hongdong Li
摘要
This paper addresses the problem of vehicle-mounted camera localization by matching a ground-level image with an overhead-view satellite map. Existing methods often treat this problem as cross-view image retrieval, and use learned deep features to match the ground-level query image to a partition (e.g., a small patch) of the satellite map. By these methods, the localization accuracy is limited by the partitioning density of the satellite map (often in the order of tens meters). Departing from the conventional wisdom of image retrieval, this paper presents a novel solution that can achieve highly-accurate localization. The key idea is to formulate the task as pose estimation and solve it by neural-net based optimization. Specifically, we design a two-branch CNN to extract robust features from the ground and satellite images, respectively. To bridge the vast cross-view domain gap, we resort to a Geometry Projection module that projects features from the satellite map to the ground-view, based on a relative camera pose. Aiming to minimize the differences between the projected features and the observed features, we employ a differentiable Levenberg-Marquardt (LM) module to search for the optimal camera pose iteratively. The entire pipeline is differentiable and runs end-to-end. Extensive experiments on standard autonomous vehicle localization datasets have confirmed the superiority of the proposed method. Notably, e.g., starting from a coarse estimate of camera location within a wide region of 40m x 40m, with an 80% likelihood our method quickly reduces the lateral location error to be within 5m on a new KITTI cross-view dataset.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper37
- Fine-Grained Cross-View Geo-Localization Using a Correlation-Aware Homography EstimatorXiaolong Wang, Runsen Xu, Zhuofan Cui, Zeyu Wan 等NeurIPS 2023 · 被引用 96 次
- Boosting 3-DoF Ground-to-Satellite Camera Localization Accuracy via Geometry-Guided Cross-View TransformerYujiao Shi, Fei Wu, Akhil Perincherry, Ankit Vora 等ICCV 2023 · 被引用 60 次
- SNAP: Self-Supervised Neural Maps for Visual Positioning and Semantic UnderstandingPaul-Edouard Sarlin, Eduard Trulls, Marc Pollefeys, Jan Hosang 等NeurIPS 2023 · 被引用 52 次
- Learning Dense Flow Field for Highly-accurate Cross-view Camera LocalizationZhenbo Song, Xianghui Ze, Jianfeng Lu, Yujiao ShiNeurIPS 2023 · 被引用 37 次
- UrBench: A Comprehensive Benchmark for Evaluating Large Multimodal Models in Multi-View Urban ScenariosBaichuan Zhou, Haote Yang, Dairong Chen, Junyan Ye 等AAAI 2025 · 被引用 34 次
它引用的顶会 Paper9
- BARF: Bundle-Adjusting Neural Radiance FieldsChen-Hsuan Lin, Wei-Chiu Ma, Antonio Torralba, Simon LuceyICCV 2021 · 被引用 867 次
- Optimal Feature Transport for Cross-View Image Geo-LocalizationYujiao Shi, Xin Yu, Liu Liu, Tong Zhang 等AAAI 2020 · 被引用 210 次
- Bridging the Domain Gap for Ground-to-Aerial Image MatchingKrishna Regmi, Mubarak ShahICCV 2019 · 被引用 191 次
- Ground-to-Aerial Image Geo-Localization With a Hard Exemplar Reweighting Triplet LossSudong Cai, Yulan Guo, Salman H. Khan, Jiwei Hu 等ICCV 2019 · 被引用 140 次
- Stochastic Attraction-Repulsion Embedding for Large Scale Image LocalizationLiu Liu, Hongdong Li, Yuchao DaiICCV 2019 · 被引用 123 次
相关 Paper
- Uncertainty-Aware Vision-Based Metric Cross-View GeolocalizationFlorian Fervers, Sebastian Bullinger, Christoph Bodensteiner, Michael Arens 等CVPR 2023
- BevSplat: Resolving Height Ambiguity via Feature-Based Gaussian Primitives for Weakly-Supervised Cross-View LocalizationQiwei Wang, Shaoxun Wu, Yujiao ShiNeurIPS 2025 · 被引用 10 次
- Where Am I Looking At? Joint Location and Orientation Estimation by Cross-View MatchingYujiao Shi, Xin Yu, Dylan Campbell, Hongdong LiCVPR 2020
- PIDLoc: Cross-View Pose Optimization Network Inspired by PID ControllersWooju Lee, Juhye Park, Dasol Hong, Changki Sung 等CVPR 2025
- Loc: Interpretable Cross-View Localization via Depth-Lifted Local Feature MatchingZimin Xia, Chenghao Xu, Alexandre AlahiICLR 2026 · 被引用 1 次
