Stochastic Attraction-Repulsion Embedding for Large Scale Image Localization
Liu Liu, Hongdong Li, Yuchao Dai
Abstract
This paper tackles the problem of large-scale imagebased localization (IBL) where the spatial location of a query image is determined by finding out the most similar reference images in a large database. For solving this problem, a critical task is to learn discriminative image representation that captures informative information relevant for localization. We propose a novel representation learning method having higher location-discriminating power. It provides the following contributions: 1) we represent a place (location) as a set of exemplar images depicting the same landmarks and aim to maximize similarities among intra-place images while minimizing similarities among inter-place images; 2) we model a similarity measure as a probability distribution on L 2 -metric distances between intra-place and inter-place image representations; 3) we propose a new Stochastic Attraction and Repulsion Embedding (SARE) loss function minimizing the KL divergence between the learned and the actual probability distributions; 4) we give theoretical comparisons between SARE, triplet ranking [2] and contrastive losses [25] . It provides insights into why SARE is better by analyzing gradients. Our SARE loss is easy to implement and pluggable to any CNN. Experiments show that our proposed method improves the localization performance on standard benchmarks by a large margin. Demonstrating the broad applicability of our method, we obtained the 3 rd place out of 209 teams in the 2018 Google Landmark Retrieval Challenge [1] . Our code and model are available at https: //github.com/Liumouliu/deepIBL .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers19
- University-1652: A Multi-view Multi-source Benchmark for Drone-based Geo-localizationZhedong Zheng, Yunchao Wei, Yi YangACM MM 2020 · 390 citations
- Rethinking Visual Geo-localization for Large-Scale ApplicationsGabriele Moreno Berton, Carlo Masone, Barbara CaputoCVPR 2022 · 235 citations
- EigenPlaces: Training Viewpoint Robust Models for Visual Place RecognitionGabriele Moreno Berton, Gabriele Trivigno, Barbara Caputo, Carlo MasoneICCV 2023 · 141 citations
- Progressive Correspondence Pruning by Consensus LearningChen Zhao, Yixiao Ge, Feng Zhu, Rui Zhao et al.ICCV 2021 · 101 citations
- Towards Seamless Adaptation of Pre-trained Models for Visual Place RecognitionFeng Lu, Lijun Zhang, Xiangyuan Lan, Shuting Dong et al.ICLR 2024 · 81 citations
Related papers
- DenserNet: Weakly Supervised Visual Localization Using Multi-Scale Feature AggregationDongfang Liu, Yiming Cui, Liqi Yan, Christos Mousas et al.AAAI 2021 · 149 citations
- Soft Contrastive Learning for Visual LocalizationJanine Thoma, Danda Pani Paudel, Luc Van GoolNeurIPS 2020 · 41 citations
- Learning Super-Features for Image RetrievalPhilippe Weinzaepfel, Thomas Lucas, Diane Larlus, Yannis KalantidisICLR 2022 · 56 citations
- Data-Efficient Large Scale Place Recognition with Graded Similarity SupervisionMaria Leyva-Vallina, Nicola Strisciuglio, Nicolai PetkovCVPR 2023
- Focus on Local: Finding Reliable Discriminative Regions for Visual Place RecognitionChangwei Wang, Shunpeng Chen, Yukun Song, Rongtao Xu et al.AAAI 2025 · 24 citations
