MTLDesc: Looking Wider to Describe Better
Changwei Wang, Rongtao Xu, Yuyang Zhang, Shibiao Xu, Weiliang Meng, Bin Fan, Xiaopeng Zhang
Abstract
Limited by the locality of convolutional neural networks, most existing local features description methods only learn local descriptors with local information and lack awareness of global and surrounding spatial context. In this work, we focus on making local descriptors ``look wider to describe better'' by learning local Descriptors with More Than Local information (MTLDesc). Specifically, we resort to context augmentation and spatial attention mechanism to make the descriptors obtain non-local awareness. First, Adaptive Global Context Augmented Module and Diverse Local Context Augmented Module are proposed to construct robust local descriptors with context information from global to local. Second, we propose the Consistent Attention Weighted Triplet Loss to leverage spatial attention awareness in both optimization and matching of local descriptors. Third, Local Features Detection with Feature Pyramid is proposed to obtain more stable and accurate keypoints localization. With the above innovations, the performance of the proposed MTLDesc significantly surpasses the current state-of-the-art local descriptors on HPatches, Aachen Day-Night localization and InLoc indoor localization benchmarks. Our code is available at https://github.com/vignywang/MTLDesc.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 21aae983-eaa1-409f-9014-88343374094dCited by top-tier papers2
- Learning Transformation-Predictive Representations for Detection and Description of Local FeaturesZihao Wang, Chunxu Wu, Yifei Yang, Zhen LiCVPR 2023
- SGPFeat: Semantic and Geometric Priors for Multi-modal Image MatchingYuxin Deng, Botian Wang, Kaining Zhang, Hao Zhang et al.AAAI 2026
Builds on4
- DISK: Learning local features with policy gradientMichal J. Tyszkiewicz, Pascal Fua, Eduard TrullsNeurIPS 2020 · 652 citations
- Key.Net: Keypoint Detection by Handcrafted and Learned CNN FiltersAxel Barroso Laguna, Edgar Riba, Daniel Ponsa, Krystian MikolajczykICCV 2019 · 323 citations
- HyNet: Learning Local Descriptor with Hybrid Similarity Measure and Triplet LossYurun Tian, Axel Barroso Laguna, Tony Ng, Vassileios Balntas et al.NeurIPS 2020 · 101 citations
- ASLFeat: Learning Local Features of Accurate Shape and LocalizationZixin Luo, Lei Zhou, Xuyang Bai, Hongkai Chen et al.CVPR 2020
Related papers
- DenserNet: Weakly Supervised Visual Localization Using Multi-Scale Feature AggregationDongfang Liu, Yiming Cui, Liqi Yan, Christos Mousas et al.AAAI 2021 · 149 citations
- Hierarchical Pyramid Diverse Attention Networks for Face RecognitionQiangchang Wang, Tianyi Wu, He Zheng, Guodong GuoCVPR 2020
- SDGMNet: Statistic-Based Dynamic Gradient Modulation for Local Descriptor LearningYuxin Deng, Jiayi MaAAAI 2024 · 13 citations
- SFD2: Semantic-Guided Feature Detection and DescriptionFei Xue, Ignas Budvytis, Roberto CipollaCVPR 2023
- D2Former: Jointly Learning Hierarchical Detectors and Contextual Descriptors via Agent-Based TransformersJianfeng He, Yuan Gao, Tianzhu Zhang, Zhe Zhang et al.CVPR 2023
