EquiPose: Exploiting Permutation Equivariance for Relative Camera Pose Estimation
Yuzhen Liu, Qiulei Dong
摘要
Relative camera pose estimation between two images is a fundamental task in 3D computer vision. Recently, many relative pose estimation networks have been explored for learning a mapping from two input images to their corresponding relative pose, however, the estimated relative poses by these methods do not have the intrinsic Pose Permutation Equivariance (PPE) property: the estimated relative pose from Image A to Image B should be the inverse of that from Image B to Image A. It means that permuting the input order of two images would cause these methods to obtain inconsistent relative poses. To address this problem, we firstly introduce the concept of PPE mapping, which indicates such a mapping that captures the intrinsic PPE property of relative poses. Then by enforcing the aforementioned PPE property, we propose a general framework for relative pose estimation, called EquiPose, which could easily accommodate various relative pose estimation networks in literature as its baseline models. We further theoretically prove that the proposed EquiPose framework could guarantee that its obtained mapping is a PPE mapping. Given a pre-trained baseline model, the proposed EquiPose framework could improve its performance even without fine-tuning, and could further boost its performance with fine-tuning. Experimental results on four public datasets demonstrate that EquiPose could significantly improve the performances of various state-of-the-art models.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper17
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- LightGlue: Local Feature Matching at Light SpeedPhilipp Lindenberger, Paul-Edouard Sarlin, Marc PollefeysICCV 2023 · 被引用 936 次
- Efficient LoFTR: Semi-Dense Local Feature Matching with Sparse-Like SpeedYifan Wang, Xingyi He, Sida Peng, Dongli Tan 等CVPR 2024 · 被引用 126 次
- HyNet: Learning Local Descriptor with Hybrid Similarity Measure and Triplet LossYurun Tian, Axel Barroso Laguna, Tony Ng, Vassileios Balntas 等NeurIPS 2020 · 被引用 101 次
- Deep Permutation Equivariant Structure from MotionDror Moran, Hodaya Koslowsky, Yoni Kasten, Haggai Maron 等ICCV 2021 · 被引用 21 次
相关 Paper
- T-Net: Effective Permutation-Equivariant Network for Two-View Correspondence LearningZhen Zhong, Guobao Xiao, Linxin Zheng, Yan Lu 等ICCV 2021 · 被引用 34 次
- RPE-PAD: Relative Pose Estimation for Pose-agnostic Anomaly DetectionZhipeng Zhang, Mengzan Qi, Rongkang Ma, Yingying Fang 等AAAI 2026
- Equicaps: Predictor-Free Pose-Aware Pre-Trained Capsule NetworksAthinoulla Konstantinou, Georgios Leontidis, Mamatha Thota, Aiden DurrantICCV 2025
- DECA: Deep viewpoint-Equivariant human pose estimation using Capsule AutoencodersNicola Garau, Niccolò Bisagno, Piotr Bródka, Nicola ConciICCV 2021 · 被引用 34 次
- PoseIRM: Enhance 3D Human Pose Estimation on Unseen Camera Settings via Invariant Risk MinimizationYanlu Cai, Weizhong Zhang, Yuan Wu, Cheng JinCVPR 2024
