MTLDesc: Looking Wider to Describe Better
Changwei Wang, Rongtao Xu, Yuyang Zhang, Shibiao Xu, Weiliang Meng, Bin Fan, Xiaopeng Zhang
摘要
Limited by the locality of convolutional neural networks, most existing local features description methods only learn local descriptors with local information and lack awareness of global and surrounding spatial context. In this work, we focus on making local descriptors ``look wider to describe better'' by learning local Descriptors with More Than Local information (MTLDesc). Specifically, we resort to context augmentation and spatial attention mechanism to make the descriptors obtain non-local awareness. First, Adaptive Global Context Augmented Module and Diverse Local Context Augmented Module are proposed to construct robust local descriptors with context information from global to local. Second, we propose the Consistent Attention Weighted Triplet Loss to leverage spatial attention awareness in both optimization and matching of local descriptors. Third, Local Features Detection with Feature Pyramid is proposed to obtain more stable and accurate keypoints localization. With the above innovations, the performance of the proposed MTLDesc significantly surpasses the current state-of-the-art local descriptors on HPatches, Aachen Day-Night localization and InLoc indoor localization benchmarks. Our code is available at https://github.com/vignywang/MTLDesc.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Learning Transformation-Predictive Representations for Detection and Description of Local FeaturesZihao Wang, Chunxu Wu, Yifei Yang, Zhen LiCVPR 2023
- SGPFeat: Semantic and Geometric Priors for Multi-modal Image MatchingYuxin Deng, Botian Wang, Kaining Zhang, Hao Zhang 等AAAI 2026
它引用的顶会 Paper4
- DISK: Learning local features with policy gradientMichal J. Tyszkiewicz, Pascal Fua, Eduard TrullsNeurIPS 2020 · 被引用 652 次
- Key.Net: Keypoint Detection by Handcrafted and Learned CNN FiltersAxel Barroso Laguna, Edgar Riba, Daniel Ponsa, Krystian MikolajczykICCV 2019 · 被引用 323 次
- HyNet: Learning Local Descriptor with Hybrid Similarity Measure and Triplet LossYurun Tian, Axel Barroso Laguna, Tony Ng, Vassileios Balntas 等NeurIPS 2020 · 被引用 101 次
- ASLFeat: Learning Local Features of Accurate Shape and LocalizationZixin Luo, Lei Zhou, Xuyang Bai, Hongkai Chen 等CVPR 2020
相关 Paper
- DenserNet: Weakly Supervised Visual Localization Using Multi-Scale Feature AggregationDongfang Liu, Yiming Cui, Liqi Yan, Christos Mousas 等AAAI 2021 · 被引用 149 次
- Hierarchical Pyramid Diverse Attention Networks for Face RecognitionQiangchang Wang, Tianyi Wu, He Zheng, Guodong GuoCVPR 2020
- SDGMNet: Statistic-Based Dynamic Gradient Modulation for Local Descriptor LearningYuxin Deng, Jiayi MaAAAI 2024 · 被引用 13 次
- SFD2: Semantic-Guided Feature Detection and DescriptionFei Xue, Ignas Budvytis, Roberto CipollaCVPR 2023
- D2Former: Jointly Learning Hierarchical Detectors and Contextual Descriptors via Agent-Based TransformersJianfeng He, Yuan Gao, Tianzhu Zhang, Zhe Zhang 等CVPR 2023
