Steerers: A Framework for Rotation Equivariant Keypoint Descriptors
Georg Bökman, Johan Edstedt, Michael Felsberg, Fredrik Kahl
Abstract
Image keypoint descriptions that are discriminative and matchable over large changes in viewpoint are vital for 3D reconstruction. However, descriptions output by learned descriptors are typically not robust to camera rotation. While they can be made more robust by, e.g., data augmentation, this degrades performance on upright images. Another approach is test-time augmentation, which incurs a significant increase in runtime. Instead, we learn a linear transform in description space that encodes rotations of the input image. We call this linear transform a steerer since it allows us to transform the descriptions as if the image was rotated. From representation theory, we know all possible steerers for the rotation group. Steerers can be optimized (A) given a fixed descriptor, (B) jointly with a descriptor or (C) we can optimize a descriptor given a fixed steerer. We perform experiments in these three settings and obtain state-of-the-art results on the rotation invariant image matching benchmarks AIMS and Roto-360. We publish code and model weights at this https url.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 823bcbda-4d2a-4e75-a015-40302596a7a3Cited by top-tier papers8
- Axis-Level Symmetry Detection with Group-Equivariant RepresentationWongyun Yu, Ahyun Seo, Minsu ChoICCV 2025 · 2 citations
- RALoc: Enhancing Outdoor LiDAR Localization via Rotation AwarenessYuyang Yang, We Li, Sheng Ao, Qingshan Xu et al.ICCV 2025 · 1 citation
- Equivariant Latent Alignment via Flow Matching under Group SymmetriesSunghyun Kim, Jaehoon Hahm, Jeongwoo Shin, Joonseok LeeICML 2026
- EquiPose: Exploiting Permutation Equivariance for Relative Camera Pose EstimationYuzhen Liu, Qiulei DongCVPR 2025
- RoMa: Robust Dense Feature MatchingJohan Edstedt, Qiyu Sun, Georg Bökman, Mårten Wadenbäck et al.CVPR 2024
Builds on13
- DISK: Learning local features with policy gradientMichal J. Tyszkiewicz, Pascal Fua, Eduard TrullsNeurIPS 2020 · 652 citations
- A Practical Method for Constructing Equivariant Multilayer Perceptrons for Arbitrary Matrix GroupsMarc Finzi, Max Welling, Andrew Gordon WilsonICML 2021 · 226 citations
- CoMIR: Contrastive Multimodal Image Representation for RegistrationNicolas Pielawski, Elisabeth Wetzer, Johan Öfverstedt, Jiahao Lu et al.NeurIPS 2020 · 110 citations
- SiLK: Simple Learned KeypointsPierre Gleize, Weiyao Wang, Matt FeiszliICCV 2023 · 87 citations
- Self-Supervised Equivariant Learning for Oriented Keypoint DetectionJongmin Lee, Byungjin Kim, Minsu ChoCVPR 2022 · 39 citations
Related papers
- Learning SO(3)-Invariant Semantic Correspondence via Local Shape TransformChunghyun Park, Seungwook Kim, Jaesik Park, Minsu ChoCVPR 2024 · 2 citations
- FILTRA: Rethinking Steerable CNN by Filter TransformBo Li, Qili Wang, Gim Hee LeeICML 2021 · 4 citations
- Steerable Transformers for Volumetric DataSoumyabrata Kundu, Risi KondorICML 2025
- Learning Rotation-Equivariant Features for Visual CorrespondenceJongmin Lee, Byungjin Kim, Seungwook Kim, Minsu ChoCVPR 2023
- Rotation-Invariant Transformer for Point Cloud MatchingHao Yu, Zheng Qin, Ji Hou, Mahdi Saleh et al.CVPR 2023
