PanoPose: Self-supervised Relative Pose Estimation for Panoramic Images
Diantao Tu, Hainan Cui, Xianwei Zheng, Shuhan Shen
摘要
Scaled relative pose estimation, i.e., estimating relative rotation and scaled relative translation between two images, has always been a major challenge in global Structure-from-Motion (SfM). This difficulty arises because the two-view relative translation computed by traditional geometric vision methods, e.g. the five-point algorithm, is scaleless. Many researchers have proposed diverse translation averaging methods to solve this problem. Instead of solving the problem in the motion averaging phase, we focus on estimating scaled relative pose with the help of panoramic cameras and deep neural networks. In this paper, a novel network, namely PanoPose, is proposed to estimate the relative motion in a fully self-supervised manner and a global SfM pipeline is built for panorama images. The proposed PanoPose comprises a depth-net and a pose-net, with self-supervision achieved by reconstructing the reference image from its neighboring images based on the estimated depth and relative pose. To maintain precise pose estimation under large viewing angle differences, we randomly rotate the panoramic images and pre-train the posenet with images before and after the rotation. To enhance scale accuracy, a fusion block is introduced to incorporate depth information into pose estimation. Extensive experiments on panoramic SfM datasets demonstrate the effectiveness of PanoPose compared with state-of-the-arts.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- PanoVGGT: Feed-Forward 3D Reconstruction from Panoramic ImageryYijing Guo, Mengjun Chao, Luo Wang, Tianyang Zhao 等CVPR 2026 · 被引用 11 次
- SoPE: Spherical Coordinate-Based Positional Embedding for Enhancing Spatial Perception of 3D LVLMsKoonting Yip, Qiyan Zhao, Wenhao Yu, Liangyu Yuan 等CVPR 2026 · 被引用 3 次
- Scene-agnostic Pose Regression for Visual LocalizationJunwei Zheng, Ruiping Liu, Yufan Chen, Zhenfang Chen 等CVPR 2025
- BADGR: Bundle Adjustment Diffusion Conditioned by Gradients for Wide-Baseline Floor Plan ReconstructionYuguang Li, Ivaylo Boyadzhiev, Zixuan Liu, Linda G. Shapiro 等CVPR 2025
- RflyPano: A Panoramic Benchmark for Ultra-low Altitude UAV Localization Powered by RflySimDun Dai, Ze Lu, Xunhua Dai, Quan QuanAAAI 2026
它引用的顶会 Paper4
- Digging Into Self-Supervised Monocular Depth EstimationClément Godard, Oisin Mac Aodha, Michael Firman, Gabriel J. BrostowICCV 2019 · 被引用 2,416 次
- LGT-Net: Indoor Panoramic Room Layout Estimation with Geometry-Aware Transformer NetworkZhigang Jiang, Zhongzheng Xiang, Jinhua Xu, Ming ZhaoCVPR 2022 · 被引用 39 次
- Improving 360 Monocular Depth Estimation via Non-local Dense Prediction Transformer and Joint Supervised and Self-Supervised LearningIlwi Yun, Hyuk-Jae Lee, Chae-Eun RheeAAAI 2022 · 被引用 34 次
- Wide-Baseline Relative Camera Pose Estimation With Directional LearningKefan Chen, Noah Snavely, Ameesh MakadiaCVPR 2021
相关 Paper
- Deep Two-View Structure-From-Motion RevisitedJianyuan Wang, Yiran Zhong, Yuchao Dai, Stan Birchfield 等CVPR 2021
- Towards Better Generalization: Joint Depth-Pose Learning Without PoseNetWang Zhao, Shaohui Liu, Yezhi Shu, Yong-Jin LiuCVPR 2020
- Calibrating Panoramic Depth Estimation for Practical Localization and MappingJunho Kim, Eun Sun Lee, Young Min KimICCV 2023 · 被引用 2 次
- Can Scale-Consistent Monocular Depth Be Learned in a Self-Supervised Scale-Invariant Manner?Lijun Wang, Yifan Wang, Linzhao Wang, Yunlong Zhan 等ICCV 2021 · 被引用 48 次
- MGSfM: Multi-Camera Geometry Driven Global Structure-from-MotionPeilin Tao, Hainan Cui, Diantao Tu, Shuhan ShenICCV 2025 · 被引用 1 次
