MFRGN: Multi-scale Feature Representation Generalization Network for Ground-to-Aerial Geo-localization
Yuntao Wang, Jinpu Zhang, Ruonan Wei, Wenbo Gao, Yuehuan Wang
摘要
Cross-area evaluation poses a significant challenge for ground-to-aerial geo-localization, in which the training and testing data are captured from entirely distinct areas. However, current methods struggle in cross-area evaluation due to their emphasis solely on learning global information from single-scale features. Some efforts alleviate this problem but rely on complex and specific technologies like pre-processing and hard sample mining. To this end, we propose a pure end-to-end solution, free from task-specific techniques, termed the Multi-scale Feature Representation Generalization Network (MFRGN) to improve generalization. Specifically, we introduce multi-scale features and explicitly utilize them by an novel global-local information representation structure with two flows, to bolster feature representations. In the global flow, we present a lightweight Self and Cross Attention Module (SCAM) to efficiently learn global embeddings. In the local flow, we develop a Global-Prompt Attention Block (GPAB) to capture discriminative features under the global embeddings as prompts. As a result, our approach generates robust descriptors representing multi-scale global and local information, thereby enhancing the model's invariance to scene variations. Extensive experiments on benchmarks show our MFRGN achieves competitive performance in same-area evaluation and improves cross-area generalization by a significant margin compared to SOTA methods. Our code is available at https://github.com/ytao-wang/MFRGN.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper2
- Video2BEV: Transforming Drone Videos to BEVs for Video-Based Geo-LocalizationHao Ju, Shaofei Huang, Si Liu, Zhedong ZhengICCV 2025 · 被引用 5 次
- Depth-Synergized Mamba Meets Memory Experts for All-Day Image Reflection SeparationSiyan Fang, Long Peng, Yuntao Wang, Ruonan Wei 等AAAI 2026 · 被引用 5 次
相关 Paper
- Aligning Geometric Spatial Layout in Cross-View Geo-Localization via Feature RecombinationQingwang Zhang, Yingying ZhuAAAI 2024 · 被引用 28 次
- GeoSURGE: Geo-localization using Semantic Fusion with Hierarchy of Geographic EmbeddingsAngel Daruna, Nicholas Meegan, Han-Pang Chiu, Supun Samarasekera 等CVPR 2026 · 被引用 2 次
- Ground-to-Aerial Image Geo-Localization With a Hard Exemplar Reweighting Triplet LossSudong Cai, Yulan Guo, Salman H. Khan, Jiwei Hu 等ICCV 2019 · 被引用 140 次
- Cross-View Geo-Localization via Learning Disentangled Geometric Layout CorrespondenceXiaohan Zhang, Xingyu Li, Waqas Sultani, Yi Zhou 等AAAI 2023 · 被引用 111 次
- Scaling Image Geo-Localization to Continent LevelPhilipp Lindenberger, Paul-Edouard Sarlin, Jan Hosang, Marc Pollefeys 等NeurIPS 2025 · 被引用 11 次
