Visual Correspondence Hallucination
Hugo Germain, Vincent Lepetit, Guillaume Bourmaud
Abstract
Given a pair of partially overlapping source and target images and a keypoint in the source image, the keypoint's correspondent in the target image can be either visible, occluded or outside the field of view. Local feature matching methods are only able to identify the correspondent's location when it is visible, while humans can also hallucinate (i.e. predict) its location when it is occluded or outside the field of view through geometric reasoning. In this paper, we bridge this gap by training a network to output a peaked probability distribution over the correspondent's location, regardless of this correspondent being visible, occluded, or outside the field of view. We experimentally demonstrate that this network is indeed able to hallucinate correspondences on pairs of images captured in scenes that were not seen at training-time. We also apply this network to an absolute camera pose estimation problem and find it is significantly more robust than state-of-the-art local feature matching-based competitors.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 6db0bc6c-5721-4e25-8b54-ed586caeedf1Cited by top-tier papers7
- ViTALiTy: Unifying Low-rank and Sparse Approximation for Vision Transformer Acceleration with a Linear Taylor AttentionJyotikrishna Dass, Shang Wu, Huihong Shi, Chaojian Li et al.HPCA 2023 · 65 citations
- Virtual Correspondence: Humans as a Cue for Extreme-View GeometryWei-Chiu Ma, Anqi Joyce Yang, Shenlong Wang, Raquel Urtasun et al.CVPR 2022 · 20 citations
- PUMP: Pyramidal and Uniqueness Matching Priors for Unsupervised Learning of Local DescriptorsJérôme Revaud, Vincent Leroy, Philippe Weinzaepfel, Boris ChidlovskiiCVPR 2022 · 16 citations
- Alligat0R: Pre-Training through Covisibility Segmentation for Relative Camera Pose RegressionThibaut Loiseau, Guillaume Bourmaud, Vincent LepetitNeurIPS 2025 · 11 citations
- Occ2Net: Robust Image Matching Based on 3D Occupancy Estimation for Occluded RegionsMiao Fan, Mingrui Chen, Chen Hu, Shuchang ZhouICCV 2023 · 7 citations
Builds on15
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou et al.ICCV 2021 · 8,921 citations
- Transformers are RNNs: Fast Autoregressive Transformers with Linear AttentionAngelos Katharopoulos, Apoorv Vyas, Nikolaos Pappas, François FleuretICML 2020 · 2,665 citations
- On the Relationship between Self-Attention and Convolutional LayersJean-Baptiste Cordonnier, Andreas Loukas, Martin JaggiICLR 2020 · 629 citations
- Learning Two-View Correspondences and Geometry Using Order-Aware NetworkJiahui Zhang, Dawei Sun, Zixin Luo, Anbang Yao et al.ICCV 2019 · 362 citations
- Neural-Guided RANSAC: Learning Where to Sample Model HypothesesEric Brachmann, Carsten RotherICCV 2019 · 282 citations
Related papers
- Scene-Aware Egocentric 3D Human Pose EstimationJian Wang, Diogo C. Luvizon, Weipeng Xu, Lingjie Liu et al.CVPR 2023
- ObjectMatch: Robust Registration using Canonical Object CorrespondencesCan Gümeli, Angela Dai, Matthias NießnerCVPR 2023
- Matching 2D Images in 3D: Metric Relative Pose from Metric CorrespondencesAxel Barroso-Laguna, Sowmya Munukutla, Victor Adrian Prisacariu, Eric BrachmannCVPR 2024
- Rethinking Correspondence-based Category-Level Object Pose EstimationHuan Ren, Wenfei Yang, Shifeng Zhang, Tianzhu ZhangCVPR 2025
- Patch2Pix: Epipolar-Guided Pixel-Level CorrespondencesQunjie Zhou, Torsten Sattler, Laura Leal-TaixéCVPR 2021
