View from Above: Orthogonal-View Aware Cross-View Localization
Shan Wang, Chuong Nguyen, Jiawei Liu, Yanhao Zhang, Sundaram Muthu, Fahira Afzal Maken, Kaihao Zhang, Hongdong Li
Abstract
This paper presents a novel aerial-to-ground feature ag-gregation strategy, tailored for the task of cross- view image-based geo-localization. Conventional vision-based methods heavily rely on matching ground-view image features with a pre-recorded image database, often through establishing planar homography correspondences via a planar ground assumption. As such, they tend to ignore features that are off-ground and not suited for handling visual occlusions, leading to unreliable localization in challenging scenarios. We propose a Top-to-Ground Aggregation (T2GA) module that capitalizes aerial orthographic views to aggregate features down to the ground level, leveraging reliable off-ground information to improve feature alignment. Furthermore, we introduce a Cycle Domain Adaptation (CycDA) loss that ensures feature extraction robustness across do-main changes. Additionally, an Equidistant Re-projection (ERP) loss is introduced to equalize the impact of all key-points on orientation error, leading to a more extended distribution of keypoints which benefits orientation estimation. On both KITTI and Ford Multi-AV datasets, our method consistently achieves the lowest mean longitudinal and lateral translations across different settings and obtains the smallest orientation error when the initial pose is less ac-curate, a more challenging setting. Further, it can complete an entire route through continual vehicle pose estimation with initial vehicle pose given only at the starting point.<sup xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">1</sup><sup xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">1</sup>Code is available at https://github.com/ShanWang-Shan/View FromAbove.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext cbabc653-5f49-4902-ab7f-1c72825eb441Cited by top-tier papers8
- BevSplat: Resolving Height Ambiguity via Feature-Based Gaussian Primitives for Weakly-Supervised Cross-View LocalizationQiwei Wang, Shaoxun Wu, Yujiao ShiNeurIPS 2025 · 10 citations
- RHO: Robust Holistic OSM-Based Metric Cross-View Geo-LocalizationJunwei Zheng, Ruize Dai, Ruiping Liu, Zichao Zeng et al.CVPR 2026 · 2 citations
- Loc: Interpretable Cross-View Localization via Depth-Lifted Local Feature MatchingZimin Xia, Chenghao Xu, Alexandre AlahiICLR 2026 · 1 citation
- Beyond Matching to Tiles: Bridging Unaligned Aerial and Satellite Views for Vision-Only UAV NavigationLiu Kejia, Haoyang Zhou, Ruoyu Xu, Peicheng Wang et al.CVPR 2026 · 1 citation
- Multi-Modal Aerial-Ground Cross-View Place Recognition with Neural ODEsSijie Wang, Rui She, Qiyu Kang, Siqi Li et al.CVPR 2025
Builds on14
- SoftTriple Loss: Deep Metric Learning Without Triplet SamplingQi Qian, Lei Shang, Baigui Sun, Juhua Hu et al.ICCV 2019 · 419 citations
- Optimal Feature Transport for Cross-View Image Geo-LocalizationYujiao Shi, Xin Yu, Liu Liu, Tong Zhang et al.AAAI 2020 · 210 citations
- TransGeo: Transformer Is All You Need for Cross-view Image Geo-localizationSijie Zhu, Mubarak Shah, Chen ChenCVPR 2022 · 189 citations
- Fine-Grained Cross-View Geo-Localization Using a Correlation-Aware Homography EstimatorXiaolong Wang, Runsen Xu, Zhuofan Cui, Zeyu Wan et al.NeurIPS 2023 · 96 citations
- Beyond Cross-view Image Retrieval: Highly Accurate Vehicle Localization Using Satellite ImageYujiao Shi, Hongdong LiCVPR 2022 · 81 citations
Related papers
- Uncertainty-Aware Vision-Based Metric Cross-View GeolocalizationFlorian Fervers, Sebastian Bullinger, Christoph Bodensteiner, Michael Arens et al.CVPR 2023
- Aligning Geometric Spatial Layout in Cross-View Geo-Localization via Feature RecombinationQingwang Zhang, Yingying ZhuAAAI 2024 · 28 citations
- Boosting 3-DoF Ground-to-Satellite Camera Localization Accuracy via Geometry-Guided Cross-View TransformerYujiao Shi, Fei Wu, Akhil Perincherry, Ankit Vora et al.ICCV 2023 · 60 citations
- Where Am I Looking At? Joint Location and Orientation Estimation by Cross-View MatchingYujiao Shi, Xin Yu, Dylan Campbell, Hongdong LiCVPR 2020
- Cross-view Geo-localization with Layer-to-Layer TransformerHongji Yang, Xiufan Lu, Yingying ZhuNeurIPS 2021 · 231 citations
