Volumetric Bundle Adjustment for Online Photorealistic Scene Capture
Ronald Clark
Abstract
Efficient photorealistic scene capture is a challenging task. Current online reconstruction systems can operate very efficiently, but images generated from the models captured by these systems are often not photorealistic. Recent approaches based on neural volume rendering can render novel views at high fidelity, but they often require a long time to train, making them impractical for applications that require real-time scene capture. In this paper, we propose a system that can reconstruct photorealistic models of complex scenes in an efficient manner. Our system processes images online, i.e. it can obtain a good quality estimate of both the scene geometry and appearance at roughly the same rate the video is captured. To achieve the efficiency, we propose a hierarchical feature volume using VDB grids. This representation is memory efficient and allows for fast querying of the scene information. Secondly, we introduce a novel optimization technique that improves the efficiency of the bundle adjustment which allows our system to converge to the target camera poses and scene geometry much faster. Experiments on real-world scenes show that our method outperforms existing systems in terms of efficiency and capture quality. To the best of our knowledge, this is the first method that can achieve online photorealistic scene capture.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers4
- HollowNeRF: Pruning Hashgrid-Based NeRFs with Trainable Collision MitigationXiufeng Xie, Riccardo Gherardi, Zhihong Pan, Stephen HuangICCV 2023 · 21 citations
- Fast Monocular Scene Reconstruction with Global-Sparse Local-Dense GridsWei Dong, Christopher B. Choy, Charles Loop, Or Litany et al.CVPR 2023
- SurfelNeRF: Neural Surfel Radiance Fields for Online Photorealistic Reconstruction of Indoor ScenesYiming Gao, Yan-Pei Cao, Ying ShanCVPR 2023
- Level-S2fM: Structure from Motion on Neural Level Set of Implicit SurfacesYuxi Xiao, Nan Xue, Tianfu Wu, Gui-Song XiaCVPR 2023
Builds on9
- Neural Sparse Voxel FieldsLingjie Liu, Jiatao Gu, Kyaw Zaw Lin, Tat-Seng Chua et al.NeurIPS 2020 · 1,535 citations
- PlenOctrees for Real-time Rendering of Neural Radiance FieldsAlex Yu, Ruilong Li, Matthew Tancik, Hao Li et al.ICCV 2021 · 1,284 citations
- MVSNeRF: Fast Generalizable Radiance Field Reconstruction from Multi-View StereoAnpei Chen, Zexiang Xu, Fuqiang Zhao, Xiaoshuai Zhang et al.ICCV 2021 · 1,024 citations
- Multi-View Stereo by Temporal Nonparametric FusionYuxin Hou, Juho Kannala, Arno SolinICCV 2019 · 99 citations
- IBRNet: Learning Multi-View Image-Based RenderingQianqian Wang, Zhicheng Wang, Kyle Genova, Pratul P. Srinivasan et al.CVPR 2021
Related papers
- BakedSDF: Meshing Neural SDFs for Real-Time View SynthesisLior Yariv, Peter Hedman, Christian Reiser, Dor Verbin et al.SIGGRAPH 2023 · 177 citations
- Neural Lumigraph RenderingPetr Kellnhofer, Lars Jebe, Andrew Jones, Ryan Spicer et al.CVPR 2021
- Compact Neural Volumetric Video Representations with Dynamic CodebooksHaoyu Guo, Sida Peng, Yunzhi Yan, Linzhan Mou et al.NeurIPS 2023 · 12 citations
- GaussianVideo: Efficient Video Representation via Hierarchical Gaussian SplattingAndrew Bond, Jui-Hsien Wang, Long Mai, Erkut Erdem et al.ICCV 2025 · 14 citations
- Plen-VDB: Memory Efficient VDB-Based Radiance Fields for Fast Training and RenderingHan Yan, Celong Liu, Chao Ma, Xing MeiCVPR 2023
