IMP: Iterative Matching and Pose Estimation with Adaptive Pooling
Fei Xue, Ignas Budvytis, Roberto Cipolla
Abstract
Previous methods solve feature matching and pose estimation using a two-stage process by first finding matches and then estimating the pose. As they ignore the geometric relationships between the two tasks, they focus on either improving the quality of matches or filtering potential outliers, leading to limited efficiency or accuracy. In contrast, we propose an iterative matching and pose estimation framework (IMP) leveraging the geometric connections between the two tasks: a few good matches are enough for a roughly accurate pose estimation; a roughly accurate pose can be used to guide the matching by providing geometric constraints. To this end, we implement a geometry-aware recurrent attention-based module which jointly outputs sparse matches and camera poses. Specifically, for each iteration, we first implicitly embed geometric information into the module via a pose-consistency loss, allowing it to predict geometry-aware matches progressively. Second, we introduce an efficient IMP, called EIMP, to dynamically discard keypoints without potential matches, avoiding redundant updating and significantly reducing the quadratic time complexity of attention computation in transformers. Experiments on YFCC100m, Scannet, and Aachen Day-Night datasets demonstrate that the proposed method outperforms previous approaches in terms of accuracy and efficiency. Code is available at https: //github.com/feixue94/imp-release
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers5
- ResMatch: Residual Attention Learning for Feature MatchingYuxin Deng, Kaining Zhang, Shihua Zhang, Yansheng Li et al.AAAI 2024 · 15 citations
- Matching While Perceiving: Enhance Image Feature Matching with Applicable Semantic AmalgamationShihua Zhang, Zhenjie Zhu, Zizhuo Li, Tao Lu et al.AAAI 2025 · 6 citations
- CoMatch: Dynamic Covisibility-Aware Transformer for Bilateral Subpixel-Level Semi-Dense Image MatchingZizhuo Li, Yifan Lu, Linfeng Tang, Shihua Zhang et al.ICCV 2025 · 1 citation
- FC-GNN: Recovering Reliable and Accurate Correspondences from InterferencesHaobo Xu, Jun Zhou, Hua Yang, Renjie Pan et al.CVPR 2024 · 1 citation
- Collaborative Feature Matching with Progressive Correspondence LearningXin Liu, Yanbing Han, Rong Qin, Bing Wang et al.AAAI 2026
Builds on17
- Training data-efficient image transformers & distillation through attentionHugo Touvron, Matthieu Cord, Matthijs Douze, Francisco Massa et al.ICML 2021 · 8,974 citations
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou et al.ICCV 2021 · 8,921 citations
- DynamicViT: Efficient Vision Transformers with Dynamic Token SparsificationYongming Rao, Wenliang Zhao, Benlin Liu, Jiwen Lu et al.NeurIPS 2021 · 1,343 citations
- Learning Two-View Correspondences and Geometry Using Order-Aware NetworkJiahui Zhang, Dawei Sun, Zixin Luo, Anbang Yao et al.ICCV 2019 · 362 citations
- Neural-Guided RANSAC: Learning Where to Sample Model HypothesesEric Brachmann, Carsten RotherICCV 2019 · 282 citations
Related papers
- End2End Multi-View Feature Matching with Differentiable Pose OptimizationBarbara Roessle, Matthias NießnerICCV 2023 · 34 citations
- 3DPCP-Net: A Lightweight Progressive 3D Correspondence Pruning Network for Accurate and Efficient Point Cloud RegistrationJingtao Wang, Zechao LiACM MM 2024 · 5 citations
- REGTR: End-to-end Point Cloud Correspondences with TransformersZi Jian Yew, Gim Hee LeeCVPR 2022 · 242 citations
- QueryPose: Sparse Multi-Person Pose Regression via Spatial-Aware Part-Level QueryYabo Xiao, Kai Su, Xiaojuan Wang, Dongdong Yu et al.NeurIPS 2022 · 32 citations
- Implicit Correspondence Learning for Image-to-Point Cloud RegistrationXinjun Li, Wenfei Yang, Jiacheng Deng, Zhixin Cheng et al.CVPR 2025
