OmniGlue: Generalizable Feature Matching with Foundation Model Guidance
Hanwen Jiang, Arjun Karpur, Bingyi Cao, Qixing Huang, André Araújo
摘要
The image matching field has been witnessing a continuous emergence of novel learnable feature matching techniques, with ever-improving performance on conventional benchmarks. However, our investigation shows that despite these gains, their potential for real-world applications is restricted by their limited generalization capabilities to novel image domains. In this paper, we introduce OmniGlue, the first learnable image matcher that is designed with generalization as a core principle. OmniGlue leverages broad knowledge from a vision foundation model to guide the feature matching process, boosting generalization to domains not seen at training time. Additionally, we propose a novel keypoint position-guided attention mechanism which disentangles spatial and appearance information, leading to enhanced matching descriptors. We perform comprehensive experiments on a suite of 7 datasets with varied image domains, including scenelevel, object-centric and aerial images. OmniGlue's novel components lead to relative gains on unseen domains of 20.9% with respect to a directly comparable reference model, while also outperforming the recent LightGlue method by 9.5% relatively. Code and model can be found at https: //hwjiang1510.github.io/OmniGlue .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper20
- EDM: Efficient Deep Feature MatchingXi Li, Tong Rao, Cihui PanICCV 2025 · 被引用 6 次
- Matching While Perceiving: Enhance Image Feature Matching with Applicable Semantic AmalgamationShihua Zhang, Zhenjie Zhu, Zizhuo Li, Tao Lu 等AAAI 2025 · 被引用 6 次
- CRFT: Consistent-Recurrent Feature Flow Transformer for Cross-Modal Image RegistrationXuecong Liu, Mengzhu Ding, Zixuan Sun, Zhang Li 等CVPR 2026 · 被引用 4 次
- TextFM: Robust Semi-dense Feature Matching with Language GuidanceZhihao Zheng, Jinglun Feng, Nirav Savaliya, Zheng-Hang Yeh 等CVPR 2026 · 被引用 2 次
- Mind the Gap: Aligning Vision Foundation Models to Image Feature MatchingYuhan Liu, Jingwen Fu, Yang Wu, Kangyi Wu 等ICCV 2025 · 被引用 2 次
它引用的顶会 Paper17
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou 等ICCV 2021 · 被引用 8,921 次
- Zero-1-to-3: Zero-shot One Image to 3D ObjectRuoshi Liu, Rundi Wu, Basile Van Hoorick, Pavel Tokmakov 等ICCV 2023 · 被引用 1,662 次
- DISK: Learning local features with policy gradientMichal J. Tyszkiewicz, Pascal Fua, Eduard TrullsNeurIPS 2020 · 被引用 652 次
- A Tale of Two Features: Stable Diffusion Complements DINO for Zero-Shot Semantic CorrespondenceJunyi Zhang, Charles Herrmann, Junhwa Hur, Luisa Polania Cabrera 等NeurIPS 2023 · 被引用 371 次
- Learning to Match Features with Seeded Graph Matching NetworkHongkai Chen, Zixin Luo, Jiahui Zhang, Lei Zhou 等ICCV 2021 · 被引用 165 次
相关 Paper
- LightGlue: Local Feature Matching at Light SpeedPhilipp Lindenberger, Paul-Edouard Sarlin, Marc PollefeysICCV 2023 · 被引用 936 次
- SuperGlue: Learning Feature Matching With Graph Neural NetworksPaul-Edouard Sarlin, Daniel DeTone, Tomasz Malisiewicz, Andrew RabinovichCVPR 2020
- Adaptive Assignment for Geometry Aware Local Feature MatchingDihe Huang, Ying Chen, Yong Liu, Jianlin Liu 等CVPR 2023
- Semantic-aware Representation Learning for Homography EstimationYuhan Liu, Qianxin Huang, Siqi Hui, Jingwen Fu 等ACM MM 2024 · 被引用 6 次
- Matcher: Segment Anything with One Shot Using All-Purpose Feature MatchingYang Liu, Muzhi Zhu, Hengtao Li, Hao Chen 等ICLR 2024 · 被引用 149 次
