Co-Attention for Conditioned Image Matching
Olivia Wiles, Sébastien Ehrhardt, Andrew Zisserman
摘要
We propose a new approach to determine correspondences between image pairs in the wild under large changes in illumination, viewpoint, context, and material. While other approaches find correspondences between pairs of images by treating the images independently, we instead condition on both images to implicitly take account of the differences between them. To achieve this, we introduce (i) a spatial attention mechanism (a co-attention module, CoAM) for conditioning the learned features on both images, and (ii) a distinctiveness score used to choose the best matches at test time. CoAM can be added to standard architectures and trained using self-supervision or supervised data, and achieves a significant performance improvement under hard conditions, e.g. large viewpoint changes. We demonstrate that models using CoAM achieve state of the art or competitive results on a wide range of tasks: local matching, camera localization, 3D reconstruction, and image stylization.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- Stacked Hybrid-Attention and Group Collaborative Learning for Unbiased Scene Graph GenerationXingning Dong, Tian Gan, Xuemeng Song, Jianlong Wu 等CVPR 2022 · 被引用 116 次
- Decoupling Makes Weakly Supervised Local Feature BetterKunhong Li, Longguang Wang, Li Liu, Qing Ran 等CVPR 2022 · 被引用 58 次
- Guide Local Feature Matching by Overlap EstimationYing Chen, Dihe Huang, Shang Xu, Jianlin Liu 等AAAI 2022 · 被引用 35 次
- Continuous Parametric Optical FlowJianqin Luo, Zhexiong Wan, Yuxin Mao, Bo Li 等NeurIPS 2023 · 被引用 6 次
- TAB: Transformer Attention Bottlenecks Enable User Intervention and Debugging in Vision-Language ModelsPooyan Rahmanzadehgervi, Hung Huy Nguyen, Rosanne Liu, Long Mai 等ICCV 2025 · 被引用 3 次
它引用的顶会 Paper3
- SuperGlue: Learning Feature Matching With Graph Neural NetworksPaul-Edouard Sarlin, Daniel DeTone, Tomasz Malisiewicz, Andrew RabinovichCVPR 2020
- MAST: A Memory-Augmented Self-Supervised TrackerZihang Lai, Erika Lu, Weidi XieCVPR 2020
- Correspondence Networks With Adaptive Neighbourhood ConsensusShuda Li, Kai Han, Theo W. Costain, Henry Howard-Jenkins 等CVPR 2020
相关 Paper
- TransforMatcher: Match-to-Match Attention for Semantic CorrespondenceSeungwook Kim, Juhong Min, Minsu ChoCVPR 2022 · 被引用 26 次
- COTR: Correspondence Transformer for Matching Across ImagesWei Jiang, Eduard Trulls, Jan Hosang, Andrea Tagliasacchi 等ICCV 2021 · 被引用 318 次
- Geometry-Free View Synthesis: Transformers and no 3D PriorsRobin Rombach, Patrick Esser, Björn OmmerICCV 2021 · 被引用 115 次
- Self-Supervised Spatial Correspondence Across ModalitiesAyush Shrivastava, Andrew OwensCVPR 2025
- LoFTR: Detector-Free Local Feature Matching With TransformersJiaming Sun, Zehong Shen, Yuang Wang, Hujun Bao 等CVPR 2021
