Adaptive Spot-Guided Transformer for Consistent Local Feature Matching
Jiahuan Yu, Jiahao Chang, Jianfeng He, Tianzhu Zhang, Jiyang Yu, Feng Wu
Abstract
Deep Space Exploration Lab, 3 China Academy of Space Technology (a)Reference (b)Linear Attention (e)Matching Result (d)Ours (c)Vanilla Attention Figure 1. The visualization of the cross attention heatmaps and matching results. We sample two similar adjacent points in the reference image (a), marked with green and red. (b) are two heatmaps of the linear cross attention in LoFTR [50] when green and red pixels are queries. (c) are two heatmaps obtained from the vanilla cross attention. (d) are two heatmaps generated by our spot-guided attention. (e) are the comparison of the final matching results produced by LoFTR [50] (top) and our method (down).
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers7
- A Consistency-Aware Spot-Guided Transformer for Versatile and Hierarchical Point Cloud RegistrationRenlang Huang, Yufan Tang, Jiming Chen, Liang LiNeurIPS 2024 · 17 citations
- HomoMatcher: Achieving Dense Feature Matching with Semi-Dense Efficiency by Homography EstimationXiaolong Wang, Lei Yu, Yingying Zhang, Jiangwei Lao et al.AAAI 2025 · 6 citations
- PRISM: PRogressive dependency maxImization for Scale-invariant image MatchingXudong Cai, Yongcai Wang, Lun Luo, Minhang Wang et al.ACM MM 2024 · 3 citations
- Learning Dense Feature Matching via Lifting Single 2D Image to 3D SpaceYingping Liang, Yutao Hu, Wenqi Shao, Ying FuICCV 2025 · 2 citations
- UniCorrn: Unified Correspondence Transformer Across 2D and 3DPrajnan Goswami, Tianye Ding, Feng Liu, Huaizu JiangCVPR 2026 · 2 citations
Builds on19
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- DISK: Learning local features with policy gradientMichal J. Tyszkiewicz, Pascal Fua, Eduard TrullsNeurIPS 2020 · 652 citations
- Learning Two-View Correspondences and Geometry Using Order-Aware NetworkJiahui Zhang, Dawei Sun, Zixin Luo, Anbang Yao et al.ICCV 2019 · 362 citations
- COTR: Correspondence Transformer for Matching Across ImagesWei Jiang, Eduard Trulls, Jan Hosang, Andrea Tagliasacchi et al.ICCV 2021 · 318 citations
- Transformer-Based Attention Networks for Continuous Pixel-Wise PredictionGuanglei Yang, Hao Tang, Mingli Ding, Nicu Sebe et al.ICCV 2021 · 246 citations
Related papers
- LoFTR: Detector-Free Local Feature Matching With TransformersJiaming Sun, Zehong Shen, Yuang Wang, Hujun Bao et al.CVPR 2021
- TopicGeo: An Efficient Unified Framework for GeolocationXin Wang, Xinlin Wang, Shuiping GouICCV 2025 · 2 citations
- ResMatch: Residual Attention Learning for Feature MatchingYuxin Deng, Kaining Zhang, Shihua Zhang, Yansheng Li et al.AAAI 2024 · 15 citations
- Cross-View Completion Models are Zero-shot Correspondence EstimatorsHonggyu An, Jin Hyeon Kim, Seonghoon Park, Jaewoo Jung et al.CVPR 2025
- Efficient LoFTR: Semi-Dense Local Feature Matching with Sparse-Like SpeedYifan Wang, Xingyi He, Sida Peng, Dongli Tan et al.CVPR 2024 · 126 citations
