Wide-Baseline Relative Camera Pose Estimation With Directional Learning
Kefan Chen, Noah Snavely, Ameesh Makadia
Abstract
Modern deep learning techniques that regress the relative camera pose between two images have difficulty dealing with challenging scenarios, such as large camera motions resulting in occlusions and significant changes in perspective that leave little overlap between images. These models continue to struggle even with the benefit of large supervised training datasets. To address the limitations of these models, we take inspiration from techniques that show regressing keypoint locations in 2D and 3D can be improved by estimating a discrete distribution over keypoint locations. Analogously, in this paper we explore improving camera pose regression by instead predicting a discrete distribution over camera poses. To realize this idea, we introduce DirectionNet, which estimates discrete distributions over the 5D relative pose space using a novel parameterization to make the estimation problem tractable. Specifically, DirectionNet factorizes relative camera pose, specified by a 3D rotation and a translation direction, into a set of 3D direction vectors. Since 3D directions can be identified with points on the sphere, Direction-Net estimates discrete distributions on the sphere as its output. We evaluate our model on challenging synthetic and real pose estimation datasets constructed from Matterport3D and InteriorNet. Promising results show a near 50% reduction in error over direct regression methods. Code will be available at https://arthurchen0518.github.io/DirectionNet.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b775e345-b6b7-4698-b70c-93828df683d3Cited by top-tier papers29
- Cameras as Rays: Pose Estimation via Ray DiffusionJason Y. Zhang, Amy Lin, Moneish Kumar, Tzu-Hsuan Yang et al.ICLR 2024 · 126 citations
- LEAP: Liberate Sparse-View 3D Modeling from Camera PosesHanwen Jiang, Zhenyu Jiang, Yue Zhao, Qixing HuangICLR 2024 · 70 citations
- Virtual Correspondence: Humans as a Cue for Extreme-View GeometryWei-Chiu Ma, Anqi Joyce Yang, Shenlong Wang, Raquel Urtasun et al.CVPR 2022 · 20 citations
- End-to-End (Instance)-Image Goal Navigation through Correspondence as an Emergent PhenomenonGuillaume Bono, Leonid Antsfeld, Boris Chidlovskii, Philippe Weinzaepfel et al.ICLR 2024 · 19 citations
- Visual Correspondence HallucinationHugo Germain, Vincent Lepetit, Guillaume BourmaudICLR 2022 · 11 citations
Builds on4
- Explaining the Ambiguity of Object Detection and 6D Pose From Visual DataFabian Manhardt, Diego Martín Arroyo, Christian Rupprecht, Benjamin Busam et al.ICCV 2019 · 139 citations
- Deep Orientation Uncertainty Learning based on a Bingham LossIgor Gilitschenski, Roshni Sahoo, Wilko Schwarting, Alexander Amini et al.ICLR 2020 · 75 citations
- D3VO: Deep Depth, Deep Pose and Deep Uncertainty for Monocular Visual OdometryNan Yang, Lukas von Stumberg, Rui Wang, Daniel CremersCVPR 2020
- SuperGlue: Learning Feature Matching With Graph Neural NetworksPaul-Edouard Sarlin, Daniel DeTone, Tomasz Malisiewicz, Andrew RabinovichCVPR 2020
Related papers
- UprightNet: Geometry-Aware Camera Orientation Estimation From Single ImagesWenqi Xian, Zhengqi Li, Noah Snavely, Matthew Fisher et al.ICCV 2019 · 52 citations
- Map-Relative Pose Regression for Visual Re-LocalizationShuai Chen, Tommaso Cavallari, Victor Adrian Prisacariu, Eric BrachmannCVPR 2024
- FAR: Flexible, Accurate and Robust 6DoF Relative Camera Pose EstimationChris Rockwell, Nilesh Kulkarni, Linyi Jin, Jeong Joon Park et al.CVPR 2024
- Uncertainty-Aware Adaptation for Self-Supervised 3D Human Pose EstimationJogendra Nath Kundu, Siddharth Seth, Pradyumna YM, Varun Jampani et al.CVPR 2022 · 41 citations
- DecenterNet: Bottom-Up Human Pose Estimation Via Decentralized Pose RepresentationTao Wang, Lei Jin, Zhang Wang, Xiaojin Fan et al.ACM MM 2023 · 14 citations
