SinGeo: Unlock Single Model's Potential for Robust Cross-View Geo-Localization
Yang Chen, Xieyuanli Chen, Junxiang Li, Jie Tang, Tao Wu
摘要
Robust cross-view geo-localization (CVGL) remains challenging despite the surge in recent progress. Existing methods still rely on field-of-view (FoV)-specific training paradigms, where models are optimized under a fixed FoV but collapse when tested on unseen FoVs and unknown orientations. This limitation necessitates deploying multiple models to cover diverse variations. Although studies have explored dynamic FoV training by simply randomizing FoVs, they failed to achieve robustness across diverse conditions-implicitly assuming all FoVs are equally difficult. To address this gap, we present SinGeo, a simple yet powerful framework that enables a Single model to realize robust cross-view Geo-localization without additional modules or explicit transformations. SinGeo employs a dual discriminative learning architecture that enhances intra-view discriminability within both ground and satellite branches, and is the first to introduce a curriculum learning strategy to achieve robust CVGL. Extensive evaluations on four benchmark datasets reveal that Sin-Geo sets state-of-the-art (SOTA) results under diverse conditions, and notably outperforms methods specifically trained for extreme FoVs. Beyond superior performance, SinGeo also exhibits cross-architecture transferability. Furthermore, we propose a consistency evaluation method to objectively assess model stability under varying views, providing an objective perspective for understanding and advancing robustness in future CVGL research. Codes are available at: https://github.com/Yangchen-nudt/SinGeo.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper11
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- A ConvNet for the 2020sZhuang Liu, Hanzi Mao, Chao-Yuan Wu, Christoph Feichtenhofer 等CVPR 2022 · 被引用 6,782 次
- University-1652: A Multi-view Multi-source Benchmark for Drone-based Geo-localizationZhedong Zheng, Yunchao Wei, Yi YangACM MM 2020 · 被引用 390 次
- Optimal Feature Transport for Cross-View Image Geo-LocalizationYujiao Shi, Xin Yu, Liu Liu, Tong Zhang 等AAAI 2020 · 被引用 210 次
- Can contrastive learning avoid shortcut solutions?Joshua Robinson, Li Sun, Ke Yu, Kayhan Batmanghelich 等NeurIPS 2021 · 被引用 185 次
相关 Paper
- PLGeo: A Patch-level Framework to Overcome Orientation Discrepancies in Cross-view Geo-localizationYiru Li, Yingying ZhuACM MM 2025 · 被引用 1 次
- First Learn, Then Review: Human-Like Continual Learning for Cross-View Geo-Localization with Limited Field of ViewLei Cheng, Daikun Liu, Zhikun Chen, Teng WangAAAI 2026
- MRGeo: Robust Cross-View Geo-Localization of Corrupted Images via Spatial and Channel Feature EnhancementLe Wu, Bo Lv, Songsong Ouyang, Yingying ZhuAAAI 2026
- MOGeo: Beyond One-to-One Cross-View Object Geo-localizationBo Lv, Qingwang Zhang, Le Wu, Yuanyuan Li 等CVPR 2026 · 被引用 2 次
- RHO: Robust Holistic OSM-Based Metric Cross-View Geo-LocalizationJunwei Zheng, Ruize Dai, Ruiping Liu, Zichao Zeng 等CVPR 2026 · 被引用 2 次
