EigenPlaces: Training Viewpoint Robust Models for Visual Place Recognition
Gabriele Moreno Berton, Gabriele Trivigno, Barbara Caputo, Carlo Masone
摘要
Visual Place Recognition is a task that aims to predict the place of an image (called query) based solely on its visual features. This is typically done through image retrieval, where the query is matched to the most similar images from a large database of geotagged photos, using learned global descriptors. A major challenge in this task is recognizing places seen from different viewpoints. To overcome this limitation, we propose a new method, called EigenPlaces, to train our neural network on images from different point of views, which embeds viewpoint robustness into the learned global descriptors. The underlying idea is to cluster the training data so as to explicitly present the model with different views of the same points of interest. The selection of this points of interest is done without the need for extra supervision. We then present experiments on the most comprehensive set of datasets in literature, finding that EigenPlaces is able to outperform previous state of the art on the majority of datasets, while requiring 60% less GPU memory for training and using 50% smaller descriptors. The code and trained models for EigenPlaces are available at https: //github.com/gmberton/EigenPlaces , while results with any other baseline can be computed with the codebase at https://github.com/gmberton/auto_VPR .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper27
- CricaVPR: Cross-Image Correlation-Aware Representation Learning for Visual Place RecognitionFeng Lu, Xiangyuan Lan, Lijun Zhang, Dongmei Jiang 等CVPR 2024 · 被引用 68 次
- SuperVLAD: Compact and Robust Image Descriptors for Visual Place RecognitionFeng Lu, Xinyao Zhang, Canming Ye, Shuting Dong 等NeurIPS 2024 · 被引用 24 次
- Focus on Local: Finding Reliable Discriminative Regions for Visual Place RecognitionChangwei Wang, Shunpeng Chen, Yukun Song, Rongtao Xu 等AAAI 2025 · 被引用 24 次
- TransLoc4D: Transformer-Based 4D Radar Place RecognitionGuohao Peng, Heshan Li, Yangyang Zhao, Jun Zhang 等CVPR 2024 · 被引用 19 次
- EMVP: Embracing Visual Foundation Model for Visual Place Recognition with Centroid-Free ProbingQibo Qiu, Shun Zhang, Haiming Gao, Honghui Yang 等NeurIPS 2024 · 被引用 13 次
它引用的顶会 Paper13
- Rethinking Visual Geo-localization for Large-Scale ApplicationsGabriele Moreno Berton, Carlo Masone, Barbara CaputoCVPR 2022 · 被引用 235 次
- TransVPR: Transformer-Based Place Recognition with Multi-Level Attention AggregationRuotong Wang, Yanqing Shen, Weiliang Zuo, Sanping Zhou 等CVPR 2022 · 被引用 167 次
- DenserNet: Weakly Supervised Visual Localization Using Multi-Scale Feature AggregationDongfang Liu, Yiming Cui, Liqi Yan, Christos Mousas 等AAAI 2021 · 被引用 149 次
- Stochastic Attraction-Repulsion Embedding for Large Scale Image LocalizationLiu Liu, Hongdong Li, Yuchao DaiICCV 2019 · 被引用 123 次
- Instance-level Image Retrieval using Reranking TransformersFuwen Tan, Jiangbo Yuan, Vicente OrdonezICCV 2021 · 被引用 116 次
相关 Paper
- MutualVPR: A Mutual Learning Framework for Resolving Supervision Inconsistencies via Adaptive ClusteringQiwen Gu, Xufei Wang, Junqiao Zhao, Siyue Tao 等NeurIPS 2025 · 被引用 6 次
- Viewpoint Invariant Dense Matching for Visual GeolocalizationGabriele Moreno Berton, Carlo Masone, Valerio Paolicelli, Barbara CaputoICCV 2021 · 被引用 48 次
- SAGE: Spatial-visual Adaptive Graph Exploration for Efficient Visual Place RecognitionShunpeng Chen, Changwei Wang, Rongtao Xu, Xingtian Pei 等ICLR 2026 · 被引用 6 次
- A Hyperdimensional One Place Signature to Represent Them All: Stackable Descriptors for Visual Place RecognitionConnor Malone, Somayeh Hussaini, Tobias Fischer, Michael MilfordICCV 2025 · 被引用 2 次
- BoQ: A Place is Worth a Bag of Learnable QueriesAmar Ali-bey, Brahim Chaib-draa, Philippe GiguèreCVPR 2024
