VolumeFusion: Deep Depth Fusion for 3D Scene Reconstruction
Jaesung Choe, Sunghoon Im, François Rameau, Minjun Kang, In So Kweon
Abstract
To reconstruct a 3D scene from a set of calibrated views, traditional multi-view stereo techniques rely on two distinct stages: local depth maps computation and global depth maps fusion. Recent studies concentrate on deep neural architectures for depth estimation by using conventional depth fusion method or direct 3D reconstruction network by regressing Truncated Signed Distance Function (TSDF). In this paper, we advocate that replicating the traditional two stages framework with deep neural networks improves both the interpretability and the accuracy of the results. As mentioned, our network operates in two steps: 1) the local computation of the local depth maps with a deep MVS technique, and, 2) the depth maps and images’ features fusion to build a single TSDF volume. In order to improve the matching performance between images acquired from very different viewpoints (e.g., large-baseline and rotations), we introduce a rotation-invariant 3D convolution kernel called PosedConv. The effectiveness of the proposed architecture is underlined via a large series of experiments conducted on the ScanNet dataset where our approach compares favorably against both traditional and deep learning techniques.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext a4b4509a-b44c-4277-b697-4684f39839c0Cited by top-tier papers22
- Point-SLAM: Dense Neural Point Cloud-based SLAMErik Sandström, Yue Li, Luc Van Gool, Martin R. OswaldICCV 2023 · 269 citations
- NeuSurf: On-Surface Priors for Neural Surface Reconstruction from Sparse Input ViewsHan Huang, Yulun Wu, Junsheng Zhou, Ge Gao et al.AAAI 2024 · 42 citations
- FatesGS: Fast and Accurate Sparse-View Surface Reconstruction Using Gaussian Splatting with Depth-Feature ConsistencyHan Huang, Yulun Wu, Chao Deng, Ge Gao et al.AAAI 2025 · 29 citations
- FineRecon: Depth-aware Feed-forward Network for Detailed 3D ReconstructionNoah Stier, Anurag Ranjan, Alex Colburn, Yajie Yan et al.ICCV 2023 · 29 citations
- Deep Point Cloud ReconstructionJaesung Choe, Byeongin Joung, François Rameau, Jaesik Park et al.ICLR 2022 · 28 citations
Builds on9
- Point-Based Multi-View Stereo NetworkRui Chen, Songfang Han, Jing Xu, Hao SuICCV 2019 · 403 citations
- TransformerFusion: Monocular RGB Scene Reconstruction using TransformersAljaz Bozic, Pablo R. Palafox, Justus Thies, Angela Dai et al.NeurIPS 2021 · 185 citations
- Multi-View Stereo by Temporal Nonparametric FusionYuxin Hou, Juho Kannala, Arno SolinICCV 2019 · 99 citations
- Normal Assisted Stereo Depth EstimationUday Kusupati, Shuo Cheng, Rui Chen, Hao SuCVPR 2020
- RoutedFusion: Learning Real-Time Depth Map FusionSilvan Weder, Johannes L. Schönberger, Marc Pollefeys, Martin R. OswaldCVPR 2020
Related papers
- MVS2D: Efficient Multiview Stereo via Attention-Driven 2D ConvolutionsZhenpei Yang, Zhile Ren, Qi Shan, Qixing HuangCVPR 2022 · 43 citations
- A Confidence-based Iterative Solver of Depths and Surface Normals for Deep Multi-view StereoWang Zhao, Shaohui Liu, Yi Wei, Hengkai Guo et al.ICCV 2021 · 16 citations
- P-MVSNet: Learning Patch-Wise Matching Confidence Aggregation for Multi-View StereoKeyang Luo, Tao Guan, Lili Ju, Haipeng Huang et al.ICCV 2019 · 254 citations
- Deep Two-View Structure-From-Motion RevisitedJianyuan Wang, Yiran Zhong, Yuchao Dai, Stan Birchfield et al.CVPR 2021
- MVSCRF: Learning Multi-View Stereo With Conditional Random FieldsYouze Xue, Jiansheng Chen, Weitao Wan, Yiqing Huang et al.ICCV 2019 · 95 citations
