FvOR: Robust Joint Shape and Pose Optimization for Few-view Object Reconstruction
Zhenpei Yang, Zhile Ren, Miguel Ángel Bautista, Zaiwei Zhang, Qi Shan, Qixing Huang
Abstract
Reconstructing an accurate 3D object model from a few image observations remains a challenging problem in computer vision. State-of-the-art approaches typically assume accurate camera poses as input, which could be difficult to obtain in realistic settings. In this paper, we present FvOR, a learning-based object reconstruction method that predicts accurate 3D models given a few images with noisy input poses. The core of our approach is a fast and robust multi-view reconstruction algorithm to jointly refine 3D geometry and camera pose estimation using learnable neural network modules. We provide a thorough benchmark of state-of-the-art approaches for this problem on ShapeNet. Our approach achieves best-in-class results. It is also two orders of magnitude faster than the recent optimization-based approach IDR [67].
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext bba382bd-d2af-4e90-9ccb-f12fd9948defCited by top-tier papers11
- PF-LRM: Pose-Free Large Reconstruction Model for Joint Pose and Shape PredictionPeng Wang, Hao Tan, Sai Bi, Yinghao Xu et al.ICLR 2024 · 170 citations
- ViP-NeRF: Visibility Prior for Sparse Input Neural Radiance FieldsNagabhushan Somraj, Rajiv SoundararajanSIGGRAPH 2023 · 38 citations
- Long-Range Grouping Transformer for Multi-View 3D ReconstructionLiying Yang, Zhenwei Zhu, Xuxin Lin, Jian Nong et al.ICCV 2023 · 11 citations
- SE(3) Equivariant Convolution and Transformer in Ray SpaceYinshuang Xu, Jiahui Lei, Kostas DaniilidisNeurIPS 2023 · 6 citations
- Emergent Extreme-View Geometry in 3D Foundation ModelsYiwen Zhang, Joseph Tung, Ruojin Cai, David Fouhey et al.CVPR 2026 · 5 citations
Builds on16
- NeuS: Learning Neural Implicit Surfaces by Volume Rendering for Multi-view ReconstructionPeng Wang, Lingjie Liu, Yuan Liu, Christian Theobalt et al.NeurIPS 2021 · 2,500 citations
- PIFu: Pixel-Aligned Implicit Function for High-Resolution Clothed Human DigitizationShunsuke Saito, Zeng Huang, Ryota Natsume, Shigeo Morishima et al.ICCV 2019 · 1,411 citations
- Multiview Neural Surface Reconstruction by Disentangling Geometry and AppearanceLior Yariv, Yoni Kasten, Dror Moran, Meirav Galun et al.NeurIPS 2020 · 1,010 citations
- UNISURF: Unifying Neural Implicit Surfaces and Radiance Fields for Multi-View ReconstructionMichael Oechsle, Songyou Peng, Andreas GeigerICCV 2021 · 885 citations
- DeepV2D: Video to Depth with Differentiable Structure from MotionZachary Teed, Jia DengICLR 2020 · 314 citations
Related papers
- SparsePose: Sparse-View Camera Pose Regression and RefinementSamarth Sinha, Jason Y. Zhang, Andrea Tagliasacchi, Igor Gilitschenski et al.CVPR 2023
- SC-NeuS: Consistent Neural Surface Reconstruction from Sparse and Noisy ViewsShi-Sheng Huang, Zi-Xin Zou, Yichi Zhang, Yan-Pei Cao et al.AAAI 2024 · 10 citations
- From Image Collections to Point Clouds With Self-Supervised Shape and Pose NetworksNavaneet K. L., Ansu Mathew, Shashank Kashyap, Wei-Chih Hung et al.CVPR 2020
- LEAP: Liberate Sparse-View 3D Modeling from Camera PosesHanwen Jiang, Zhenyu Jiang, Yue Zhao, Qixing HuangICLR 2024 · 70 citations
- H3D-Net: Few-Shot High-Fidelity 3D Head ReconstructionEduard Ramon, Gil Triginer, Janna Escur, Albert Pumarola et al.ICCV 2021 · 108 citations
