SparsePose: Sparse-View Camera Pose Regression and Refinement
Samarth Sinha, Jason Y. Zhang, Andrea Tagliasacchi, Igor Gilitschenski, David B. Lindell
Abstract
Camera pose estimation is a key step in standard 3D reconstruction pipelines that operate on a dense set of images of a single object or scene. However, methods for pose estimation often fail when only a few images are available because they rely on the ability to robustly identify and match visual features between image pairs. While these methods can work robustly with dense camera views, capturing a large set of images can be time-consuming or impractical. We propose SparsePose for recovering accurate camera poses given a sparse set of wide-baseline images (fewer than 10). The method learns to regress initial camera poses and then iteratively refine them after training on a large-scale dataset of objects (Co3D: Common Objects in 3D). SparsePose significantly outperforms conventional and learning-based baselines in recovering accurate camera rotations and translations. We also demonstrate our pipeline for high-fidelity 3D reconstruction using only 5-9 images of an object. Project webpage: https: //sparsepose.github.io/
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers24
- PF-LRM: Pose-Free Large Reconstruction Model for Joint Pose and Shape PredictionPeng Wang, Hao Tan, Sai Bi, Yinghao Xu et al.ICLR 2024 · 170 citations
- PoseDiffusion: Solving Pose Estimation via Diffusion-aided Bundle AdjustmentJianyuan Wang, Christian Rupprecht, David NovotnýICCV 2023 · 158 citations
- Cameras as Rays: Pose Estimation via Ray DiffusionJason Y. Zhang, Amy Lin, Moneish Kumar, Tzu-Hsuan Yang et al.ICLR 2024 · 126 citations
- LEAP: Liberate Sparse-View 3D Modeling from Camera PosesHanwen Jiang, Zhenyu Jiang, Yue Zhao, Qixing HuangICLR 2024 · 70 citations
- VGGSfM: Visual Geometry Grounded Deep Structure from MotionJianyuan Wang, Nikita Karaev, Christian Rupprecht, David NovotnýCVPR 2024 · 48 citations
Builds on35
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou et al.ICCV 2021 · 8,921 citations
- Mip-NeRF 360: Unbounded Anti-Aliased Neural Radiance FieldsJonathan T. Barron, Ben Mildenhall, Dor Verbin, Pratul P. Srinivasan et al.CVPR 2022 · 1,603 citations
- Depth-supervised NeRF: Fewer Views and Faster Training for FreeKangle Deng, Andrew Liu, Jun-Yan Zhu, Deva RamananCVPR 2022 · 756 citations
- Block-NeRF: Scalable Large Scene Neural View SynthesisMatthew Tancik, Vincent Casser, Xinchen Yan, Sabeek Pradhan et al.CVPR 2022 · 702 citations
Related papers
- FLARE: Feed-forward Geometry, Appearance and Camera Estimation from Uncalibrated Sparse ViewsShangzhan Zhang, Jianyuan Wang, Yinghao Xu, Nan Xue et al.CVPR 2025
- FvOR: Robust Joint Shape and Pose Optimization for Few-view Object ReconstructionZhenpei Yang, Zhile Ren, Miguel Ángel Bautista, Zaiwei Zhang et al.CVPR 2022 · 18 citations
- Reconstruct Locally, Localize Globally: A Model Free Method for Object Pose EstimationMing Cai, Ian ReidCVPR 2020
- Sparse-view Pose Estimation and Reconstruction via Analysis by Generative SynthesisQitao Zhao, Shubham TulsianiNeurIPS 2024 · 10 citations
- FS6D: Few-Shot 6D Pose Estimation of Novel ObjectsYisheng He, Yao Wang, Haoqiang Fan, Jian Sun et al.CVPR 2022 · 85 citations
