ELF: Embedded Localisation of Features in Pre-Trained CNN
Assia Benbihi, Matthieu Geist, Cédric Pradalier
Abstract
This paper introduces a novel feature detector based only on information embedded inside a CNN trained on standard tasks (e.g. classification). While previous works already show that the features of a trained CNN are suitable descriptors, we show here how to extract the feature locations from the network to build a detector. This information is computed from the gradient of the feature map with respect to the input image. This provides a saliency map with local maxima on relevant keypoint locations. Contrary to recent CNN-based detectors, this method requires neither supervised training nor finetuning. We evaluate how repeatable and how ‘matchable’ the detected keypoints are with the repeatability and matching scores. Matchability is measured with a simple descriptor introduced for the sake of the evaluation. This novel detector reaches similar performances on the standard evaluation HPatches dataset, as well as comparable robustness against illumination and viewpoint changes on Webcam and photo-tourism images. These results show that a CNN trained on a standard task embeds feature location information that is as relevant as when the CNN is specifically trained for feature detection.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 900a0b85-64fb-4ee3-8160-4a726ffdf607Cited by top-tier papers5
- Reinforced Feature Points: Optimizing Feature Detection and Description for a High-Level TaskAritra Bhowmik, Stefan Gumhold, Carsten Rother, Eric BrachmannCVPR 2020
- Neural Reprojection Error: Merging Feature Learning and Camera Pose EstimationHugo Germain, Vincent Lepetit, Guillaume BourmaudCVPR 2021
- D2Former: Jointly Learning Hierarchical Detectors and Contextual Descriptors via Agent-Based TransformersJianfeng He, Yuan Gao, Tianzhu Zhang, Zhe Zhang et al.CVPR 2023
- Deep Lucas-Kanade Homography for Multimodal Image AlignmentYiming Zhao, Xinming Huang, Ziming ZhangCVPR 2021
- Wide-Baseline Multi-Camera Calibration Using Person Re-IdentificationYan Xu, Yu-Jhe Li, Xinshuo Weng, Kris KitaniCVPR 2021
Related papers
- Key.Net: Keypoint Detection by Handcrafted and Learned CNN FiltersAxel Barroso Laguna, Edgar Riba, Daniel Ponsa, Krystian MikolajczykICCV 2019 · 323 citations
- Self-Supervised Equivariant Learning for Oriented Keypoint DetectionJongmin Lee, Byungjin Kim, Minsu ChoCVPR 2022 · 39 citations
- GLAMpoints: Greedily Learned Accurate Match PointsPrune Truong, Stefanos Apostolopoulos, Agata Mosinska, Samuel Stucky et al.ICCV 2019 · 77 citations
- DenserNet: Weakly Supervised Visual Localization Using Multi-Scale Feature AggregationDongfang Liu, Yiming Cui, Liqi Yan, Christos Mousas et al.AAAI 2021 · 149 citations
- SOLD2: Self-Supervised Occlusion-Aware Line Description and DetectionRémi Pautrat, Juan-Ting Lin, Viktor Larsson, Martin R. Oswald et al.CVPR 2021
