SinGeo: Unlock Single Model's Potential for Robust Cross-View Geo-Localization
Yang Chen, Xieyuanli Chen, Junxiang Li, Jie Tang, Tao Wu
Abstract
Robust cross-view geo-localization (CVGL) remains challenging despite the surge in recent progress. Existing methods still rely on field-of-view (FoV)-specific training paradigms, where models are optimized under a fixed FoV but collapse when tested on unseen FoVs and unknown orientations. This limitation necessitates deploying multiple models to cover diverse variations. Although studies have explored dynamic FoV training by simply randomizing FoVs, they failed to achieve robustness across diverse conditions-implicitly assuming all FoVs are equally difficult. To address this gap, we present SinGeo, a simple yet powerful framework that enables a Single model to realize robust cross-view Geo-localization without additional modules or explicit transformations. SinGeo employs a dual discriminative learning architecture that enhances intra-view discriminability within both ground and satellite branches, and is the first to introduce a curriculum learning strategy to achieve robust CVGL. Extensive evaluations on four benchmark datasets reveal that Sin-Geo sets state-of-the-art (SOTA) results under diverse conditions, and notably outperforms methods specifically trained for extreme FoVs. Beyond superior performance, SinGeo also exhibits cross-architecture transferability. Furthermore, we propose a consistency evaluation method to objectively assess model stability under varying views, providing an objective perspective for understanding and advancing robustness in future CVGL research. Codes are available at: https://github.com/Yangchen-nudt/SinGeo.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on11
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- A ConvNet for the 2020sZhuang Liu, Hanzi Mao, Chao-Yuan Wu, Christoph Feichtenhofer et al.CVPR 2022 · 6,782 citations
- University-1652: A Multi-view Multi-source Benchmark for Drone-based Geo-localizationZhedong Zheng, Yunchao Wei, Yi YangACM MM 2020 · 390 citations
- Optimal Feature Transport for Cross-View Image Geo-LocalizationYujiao Shi, Xin Yu, Liu Liu, Tong Zhang et al.AAAI 2020 · 210 citations
- Can contrastive learning avoid shortcut solutions?Joshua Robinson, Li Sun, Ke Yu, Kayhan Batmanghelich et al.NeurIPS 2021 · 185 citations
Related papers
- PLGeo: A Patch-level Framework to Overcome Orientation Discrepancies in Cross-view Geo-localizationYiru Li, Yingying ZhuACM MM 2025 · 1 citation
- First Learn, Then Review: Human-Like Continual Learning for Cross-View Geo-Localization with Limited Field of ViewLei Cheng, Daikun Liu, Zhikun Chen, Teng WangAAAI 2026
- MRGeo: Robust Cross-View Geo-Localization of Corrupted Images via Spatial and Channel Feature EnhancementLe Wu, Bo Lv, Songsong Ouyang, Yingying ZhuAAAI 2026
- MOGeo: Beyond One-to-One Cross-View Object Geo-localizationBo Lv, Qingwang Zhang, Le Wu, Yuanyuan Li et al.CVPR 2026 · 2 citations
- RHO: Robust Holistic OSM-Based Metric Cross-View Geo-LocalizationJunwei Zheng, Ruize Dai, Ruiping Liu, Zichao Zeng et al.CVPR 2026 · 2 citations
