PUMP: Pyramidal and Uniqueness Matching Priors for Unsupervised Learning of Local Descriptors
Jérôme Revaud, Vincent Leroy, Philippe Weinzaepfel, Boris Chidlovskii
Abstract
Existing approaches for learning local image descriptors have shown remarkable achievements in a wide range of geometric tasks. However, most of them require perpixel correspondence-level supervision, which is difficult to acquire at scale and in high quality. In this paper, we propose to explicitly integrate two matching priors in a single loss in order to learn local descriptors without supervision. Given two images depicting the same scene, we extract pixel descriptors and build a correlation volume. The first prior enforces the local consistency of matches in this volume via a pyramidal structure iteratively constructed using a non-parametric module. The second prior exploits the fact that each descriptor should match with at most one descriptor from the other image. We combine our unsupervised loss with a standard self-supervised loss trained from synthetic image augmentations. Feature descriptors learned by the proposed approach outperform their fully- and self-supervised counterparts on various geometric benchmarks such as visual localization and image matching, achieving state-of-the-art performance. Project webpage: https://europe.naverlabs.com/research/3d-vision/pump.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 7c76bcd9-fabd-4bf4-8726-96ffbaf9fb65Cited by top-tier papers4
- Unifying Feature and Cost Aggregation with Transformers for Semantic and Visual CorrespondenceSunghwan Hong, Seokju Cho, Seungryong Kim, Stephen LinICLR 2024 · 16 citations
- FC-GNN: Recovering Reliable and Accurate Correspondences from InterferencesHaobo Xu, Jun Zhou, Hua Yang, Renjie Pan et al.CVPR 2024 · 1 citation
- SFD2: Semantic-Guided Feature Detection and DescriptionFei Xue, Ignas Budvytis, Roberto CipollaCVPR 2023
- Learning Rotation-Equivariant Features for Visual CorrespondenceJongmin Lee, Byungjin Kim, Seungwook Kim, Minsu ChoCVPR 2023
Builds on24
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Learning With Average Precision: Training Image Retrieval With a Listwise LossJérôme Revaud, Jon Almazán, Rafael S. Rezende, César Roberto de SouzaICCV 2019 · 424 citations
- COTR: Correspondence Transformer for Matching Across ImagesWei Jiang, Eduard Trulls, Jan Hosang, Andrea Tagliasacchi et al.ICCV 2021 · 318 citations
- Dual-Resolution Correspondence NetworksXinghui Li, Kai Han, Shuda Li, Victor PrisacariuNeurIPS 2020 · 207 citations
- Learning to Match Features with Seeded Graph Matching NetworkHongkai Chen, Zixin Luo, Jiahui Zhang, Lei Zhou et al.ICCV 2021 · 165 citations
Related papers
- Learning Transformation-Predictive Representations for Detection and Description of Local FeaturesZihao Wang, Chunxu Wu, Yifei Yang, Zhen LiCVPR 2023
- Focus on Local: Finding Reliable Discriminative Regions for Visual Place RecognitionChangwei Wang, Shunpeng Chen, Yukun Song, Rongtao Xu et al.AAAI 2025 · 24 citations
- Unsupervised Learning of Landmarks by Descriptor Vector ExchangeJames Thewlis, Samuel Albanie, Hakan Bilen, Andrea VedaldiICCV 2019 · 70 citations
- Self-Supervised Image Representation Learning with Geometric Set ConsistencyNenglun Chen, Lei Chu, Hao Pan, Yan Lu et al.CVPR 2022 · 8 citations
- Self-Supervised Dense Consistency Regularization for Image-to-Image TranslationMinsu Ko, Eunju Cha, Sungjoo Suh, Huijin Lee et al.CVPR 2022 · 25 citations
