Learning Soft Estimator of Keypoint Scale and Orientation with Probabilistic Covariant Loss
Pei Yan, Yihua Tan, Shengzhou Xiong, Yuan Tai, Yansheng Li
Abstract
Estimating keypoint scale and orientation is crucial to extracting invariant features under significant geometric changes. Recently, the estimators based on self-supervised learning have been designed to adapt to complex imaging conditions. Such learning-based estimators generally predict a single scalar for the keypoint scale or orientation, called hard estimators. However, hard estimators are difficult to handle the local patches containing structures of different objects or multiple edges. In this paper, a Soft Self-Supervised Estimator (S3Esti) is proposed to overcome this problem by learning to predict multiple scales and orientations. S3Esti involves three core factors. First, the estimator is constructed to predict the discrete distributions of scales and orientations. The elements with high confidence will be kept as the final scales and orientations. Second, a probabilistic covariant loss is proposed to improve the consistency of the scale and orientation distributions under different transformations. Third, an optimization algorithm is designed to minimize the loss function, whose convergence is proved in theory. When combined with different keypoint extraction models, S3Esti generally improves over 50% accuracy in image matching tasks under significant viewpoint changes. In the 3D reconstruction task, S3Esti decreases more than 10% reprojection error and improves the number of registered images. [code release]
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 708e3483-8d3d-4110-a046-c6d93182cd5fCited by top-tier papers4
- ResMatch: Residual Attention Learning for Feature MatchingYuxin Deng, Kaining Zhang, Shihua Zhang, Yansheng Li et al.AAAI 2024 · 15 citations
- HOMO-Feature: Cross-Arbitrary-Modal Image Matching with Homomorphism of Organized Major OrientationChenzhong Gao, Wei Li, Desheng WengICCV 2025 · 3 citations
- Learning Affine Correspondences by Integrating Geometric ConstraintsPengju Sun, Banglei Guan, Zhenbao Yu, Yang Shang et al.CVPR 2025
- Learning Rotation-Equivariant Features for Visual CorrespondenceJongmin Lee, Byungjin Kim, Seungwook Kim, Minsu ChoCVPR 2023
Builds on5
- Fully Convolutional Geometric FeaturesChristopher B. Choy, Jaesik Park, Vladlen KoltunICCV 2019 · 807 citations
- Pose Correction for Highly Accurate Visual Localization in Large-scale Indoor SpacesJanghun Hyeon, Joohyung Kim, Nakju Lett DohICCV 2021 · 26 citations
- ASLFeat: Learning Local Features of Accurate Shape and LocalizationZixin Luo, Lei Zhou, Xuyang Bai, Hongkai Chen et al.CVPR 2020
- Wide-Baseline Relative Camera Pose Estimation With Directional LearningKefan Chen, Noah Snavely, Ameesh MakadiaCVPR 2021
- LoFTR: Detector-Free Local Feature Matching With TransformersJiaming Sun, Zehong Shen, Yuang Wang, Hujun Bao et al.CVPR 2021
Related papers
- Self-Supervised Equivariant Learning for Oriented Keypoint DetectionJongmin Lee, Byungjin Kim, Minsu ChoCVPR 2022 · 39 citations
- Learning SO(3)-Invariant Semantic Correspondence via Local Shape TransformChunghyun Park, Seungwook Kim, Jaesik Park, Minsu ChoCVPR 2024 · 2 citations
- SiLK: Simple Learned KeypointsPierre Gleize, Weiyao Wang, Matt FeiszliICCV 2023 · 87 citations
- ScaleNet: A Shallow Architecture for Scale EstimationAxel Barroso Laguna, Yurun Tian, Krystian MikolajczykCVPR 2022
- Learning Transformation-Predictive Representations for Detection and Description of Local FeaturesZihao Wang, Chunxu Wu, Yifei Yang, Zhen LiCVPR 2023
