FAR: Flexible, Accurate and Robust 6DoF Relative Camera Pose Estimation
Chris Rockwell, Nilesh Kulkarni, Linyi Jin, Jeong Joon Park, Justin Johnson, David F. Fouhey
Abstract
Estimating relative camera poses between images has been a central problem in computer vision. Methods that find correspondences and solve for the fundamental matrix offer high precision in most cases. Conversely, methods predicting pose directly using neural networks are more robust to limited overlap and can infer absolute translation scale, but at the expense of reduced precision. We show how to combine the best of both methods; our approach yields results that are both precise and robust, while also accurately inferring translation scales. At the heart of our model lies a Transformer that (1) learns to balance between solved and learned pose estimations, and (2) provides a prior to guide a solver. A comprehensive analysis supports our design choices and demonstrates that our method adapts flexibly to various feature extractors and correspondence estimators, showing state-of-the-art performance in 6DoF pose estimation on Matterport3D, Inte-riorNet, StreetLearn, and Map-free Relocalization. Project
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 5166114d-2ca3-4389-86a2-930039d73a9dCited by top-tier papers10
- DUSt3R: Geometric 3D Vision Made EasyShuzhe Wang, Vincent Leroy, Yohann Cabon, Boris Chidlovskii et al.CVPR 2024 · 302 citations
- FLARE: Feed-forward Geometry, Appearance and Camera Estimation from Uncalibrated Sparse ViewsShangzhan Zhang, Jianyuan Wang, Yinghao Xu, Nan Xue et al.CVPR 2025
- Scene-agnostic Pose Regression for Visual LocalizationJunwei Zheng, Ruiping Liu, Yufan Chen, Zhenfang Chen et al.CVPR 2025
- Unifying Correspondence, Pose and NeRF for Generalized Pose-Free Novel View SynthesisSunghwan Hong, Jaewoo Jung, Heeseong Shin, Jiaolong Yang et al.CVPR 2024
- EquiPose: Exploiting Permutation Equivariance for Relative Camera Pose EstimationYuzhen Liu, Qiulei DongCVPR 2025
Builds on23
- Vision Transformers for Dense PredictionRené Ranftl, Alexey Bochkovskiy, Vladlen KoltunICCV 2021 · 2,647 citations
- Habitat: A Platform for Embodied AI ResearchManolis Savva, Jitendra Malik, Devi Parikh, Dhruv Batra et al.ICCV 2019 · 1,863 citations
- DROID-SLAM: Deep Visual SLAM for Monocular, Stereo, and RGB-D CamerasZachary Teed, Jia DengNeurIPS 2021 · 1,248 citations
- LightGlue: Local Feature Matching at Light SpeedPhilipp Lindenberger, Paul-Edouard Sarlin, Marc PollefeysICCV 2023 · 936 citations
- BARF: Bundle-Adjusting Neural Radiance FieldsChen-Hsuan Lin, Wei-Chiu Ma, Antonio Torralba, Simon LuceyICCV 2021 · 867 citations
Related papers
- Relative Pose Estimation through Affine Corrections of Monocular Depth PriorsYifan Yu, Shaohui Liu, Rémi Pautrat, Marc Pollefeys et al.CVPR 2025
- Improving Transformer-based Image Matching by Cascaded Capturing Spatially Informative KeypointsChenjie Cao, Yanwei FuICCV 2023 · 23 citations
- Back to the Feature: Learning Robust Camera Localization From Pixels To PosePaul-Edouard Sarlin, Ajaykumar Unagar, Måns Larsson, Hugo Germain et al.CVPR 2021
- Learning Multi-Scene Absolute Pose Regression with TransformersYoli Shavit, Ron Ferens, Yosi KellerICCV 2021 · 163 citations
- RayPose: Ray Bundling Diffusion for Template Views in Unseen 6D Object Pose EstimationJunwen Huang, Shishir Reddy Vutukur, Peter KT Yu, Nassir Navab et al.ICCV 2025 · 1 citation
