SGAT: Learning Feature Matching with Singularity-enhanced Graph Attention Network
Yizhuo Zhang, Kun Sun, Chang Tang, Yuanyuan Liu, Xin Li
Abstract
The task of image feature matching aims to establish correct correspondences between images from two different views. While approaches based on attention mechanisms have demonstrated remarkable advancements in image feature matching, they still encounter substantial limitations. Specifically, current graph attention network approaches face performance bottlenecks in complex scenarios, such as lowtexture regions or occlusions. This limitation stems from the self-attention mechanism, which, when lacking effective guidance, can lead to divergent attention weights or incorrect focus on regions with low discriminability, resulting in matching failures in low-texture environments. Inspired by how humans focus on distinctive regions when performing cross-view matching, we enhance attention to singular points in images that are salient, unique and have high cross-view matching potential during information aggregation, thereby improving matching capability. To realize the aforementioned strategies, we develop a novel Singularity-enhanced Graph Attention Network (SGAT). SGAT leverages Co-potentiality and Multi-Scale Singularity as prior guidance, and designs a Singularity-aware Attention mechanism and a Co-potentiality Guided Attention mechanism, specifically enhancing the perception of singularity and matching potential during feature interaction. Experimental results on multiple datasets, including ScanNet1500, demonstrate that our method outperforms current state-of-the-art sparse matching methods. In particular, the improvement is most pronounced in complex scenarios such as low-texture environments, significantly enhancing the accuracy and robustness of image matching and its downstream tasks.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on24
- LightGlue: Local Feature Matching at Light SpeedPhilipp Lindenberger, Paul-Edouard Sarlin, Marc PollefeysICCV 2023 · 936 citations
- ScanNet++: A High-Fidelity Dataset of 3D Indoor ScenesChandan Yeshwanth, Yueh-Cheng Liu, Matthias Nießner, Angela DaiICCV 2023 · 659 citations
- DISK: Learning local features with policy gradientMichal J. Tyszkiewicz, Pascal Fua, Eduard TrullsNeurIPS 2020 · 652 citations
- Key.Net: Keypoint Detection by Handcrafted and Learned CNN FiltersAxel Barroso Laguna, Edgar Riba, Daniel Ponsa, Krystian MikolajczykICCV 2019 · 323 citations
- Quadtree Attention for Vision TransformersShitao Tang, Jiahui Zhang, Siyu Zhu, Ping TanICLR 2022 · 194 citations
Related papers
- End2End Multi-View Feature Matching with Differentiable Pose OptimizationBarbara Roessle, Matthias NießnerICCV 2023 · 34 citations
- A Structured Graph Attention Network for Vehicle Re-IdentificationYangchun Zhu, Zheng-Jun Zha, Tianzhu Zhang, Jiawei Liu et al.ACM MM 2020 · 39 citations
- ResMatch: Residual Attention Learning for Feature MatchingYuxin Deng, Kaining Zhang, Shihua Zhang, Yansheng Li et al.AAAI 2024 · 15 citations
- Scene-Aware Feature MatchingXiaoyong Lu, Yaping Yan, Tong Wei, Songlin DuICCV 2023 · 9 citations
- Learning Intra-View and Cross-View Geometric Knowledge for Stereo MatchingRui Gong, Weide Liu, Zaiwang Gu, Xulei Yang et al.CVPR 2024
