University-1652: A Multi-view Multi-source Benchmark for Drone-based Geo-localization
Zhedong Zheng, Yunchao Wei, Yi Yang
Abstract
We consider the problem of cross-view geo-localization. The primary challenge is to learn the robust feature against large viewpoint changes. Existing benchmarks can help, but are limited in the number of viewpoints. Image pairs, containing two viewpoints, e.g., satellite and ground, are usually provided, which may compromise the feature learning. Besides phone cameras and satellites, in this paper, we argue that drones could serve as the third platform to deal with the geo-localization problem. In contrast to traditional ground-view images, drone-view images meet fewer obstacles, e.g., trees, and provide a comprehensive view when flying around the target place. To verify the effectiveness of the drone platform, we introduce a new multi-view multi-source benchmark for drone-based geo-localization, named University-1652. University-1652 contains data from three platforms, i.e., synthetic drones, satellites and ground cameras of 1,652 university buildings around the world. To our knowledge, University-1652 is the first drone-based geo-localization dataset and enables two new tasks, i.e., drone-view target localization and drone navigation. As the name implies, drone-view target localization intends to predict the location of the target place via drone-view images. On the other hand, given a satellite-view query image, drone navigation is to drive the drone to the area of interest in the query. We use this dataset to analyze a variety of off-the-shelf CNN features and propose a strong CNN baseline on this challenging dataset. The experiments show that University-1652 helps the model to learn viewpoint-invariant features and also has good generalization ability in real-world scenarios.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 7a2862e5-31d7-4939-a7ac-b232c0a5d25eCited by top-tier papers37
- Cross-view Geo-localization with Layer-to-Layer TransformerHongji Yang, Xiufan Lu, Yingying ZhuNeurIPS 2021 · 231 citations
- Sample4Geo: Hard Negative Sampling For Cross-View Geo-LocalisationFabian Deuser, Konrad Habel, Norbert OswaldICCV 2023 · 161 citations
- Cross-View Geo-Localization via Learning Disentangled Geometric Layout CorrespondenceXiaohan Zhang, Xingyu Li, Waqas Sultani, Yi Zhou et al.AAAI 2023 · 111 citations
- Composed Image Retrieval with Text Feedback via Multi-grained Uncertainty RegularizationYiyang Chen, Zhedong Zheng, Wei Ji, Leigang Qu et al.ICLR 2024 · 80 citations
- Multi-View Consistent Generative Adversarial Networks for 3D-aware Image SynthesisXuanmeng Zhang, Zhedong Zheng, Daiheng Gao, Bang Zhang et al.CVPR 2022 · 37 citations
Builds on4
- Bridging the Domain Gap for Ground-to-Aerial Image MatchingKrishna Regmi, Mubarak ShahICCV 2019 · 191 citations
- Ground-to-Aerial Image Geo-Localization With a Hard Exemplar Reweighting Triplet LossSudong Cai, Yulan Guo, Salman H. Khan, Jiwei Hu et al.ICCV 2019 · 140 citations
- Stochastic Attraction-Repulsion Embedding for Large Scale Image LocalizationLiu Liu, Hongdong Li, Yuchao DaiICCV 2019 · 123 citations
- Meta Parsing Networks: Towards Generalized Few-shot Scene Parsing with Adaptive Metric LearningPeike Li, Yunchao Wei, Yi YangACM MM 2020 · 22 citations
Related papers
- UniGeoRS: A Unified Benchmark for Tri-view Geo-LocalizationXiao Liang, Huaizhi Tang, Feiyang Zhang, Shiji Yuan et al.CVPR 2026
- Game4Loc: A UAV Geo-Localization Benchmark from Game DataYuxiang Ji, Boyong He, Zhuoyue Tan, Liaoni WuAAAI 2025 · 35 citations
- Video2BEV: Transforming Drone Videos to BEVs for Video-Based Geo-LocalizationHao Ju, Shaofei Huang, Si Liu, Zhedong ZhengICCV 2025 · 5 citations
- MMGeo: Multimodal Compositional Geo-Localization for UAVsYuxiang Ji, Boyong He, Zhuoyue Tan, Liaoni WuICCV 2025 · 5 citations
- From Coarse to Fine: A Matching and Alignment Framework for Unsupervised Cross-View Geo-LocalizationXueyi Wang, Lele Zhang, Zheng Fan, Yang Liu et al.AAAI 2025 · 12 citations
