Learning Neural Implicit through Volume Rendering with Attentive Depth Fusion Priors
Pengchong Hu, Zhizhong Han
Abstract
Learning neural implicit representations has achieved remarkable performance in 3D reconstruction from multi-view images. Current methods use volume rendering to render implicit representations into either RGB or depth images that are supervised by multi-view ground truth. However, rendering a view each time suffers from incomplete depth at holes and unawareness of occluded structures from the depth supervision, which severely affects the accuracy of geometry inference via volume rendering. To resolve this issue, we propose to learn neural implicit representations from multi-view RGBD images through volume rendering with an attentive depth fusion prior. Our prior allows neural networks to perceive coarse 3D structures from the Truncated Signed Distance Function (TSDF) fused from all depth images available for rendering. The TSDF enables accessing the missing depth at holes on one depth image and the occluded parts that are invisible from the current view. By introducing a novel attention mechanism, we allow neural networks to directly use the depth fusion prior with the inferred occupancy as the learned implicit function. Our attention mechanism works with either a one-time fused TSDF that represents a whole scene or an incrementally fused TSDF that represents a partial scene in the context of Simultaneous Localization and Mapping (SLAM). Our evaluations on widely used benchmarks including synthetic and real-world scans show our superiority over the latest neural implicit methods. Please see our project page for code and data at https://machineperceptionlab.github.io/Attentive_DF_Prior/ .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext bfee57db-8b8c-4dba-b5d6-1bae118034dcCited by top-tier papers10
- Neural Signed Distance Function Inference through Splatting 3D Gaussians Pulled on Zero-Level SetWenyuan Zhang, Yu-Shen Liu, Zhizhong HanNeurIPS 2024 · 58 citations
- MultiPull: Detailing Signed Distance Functions by Pulling Multi-Level Queries at Multi-StepTakeshi Noda, Chao Chen, Weiqi Zhang, Xinhai Liu et al.NeurIPS 2024 · 19 citations
- Inferring Neural Signed Distance Functions by Overfitting on Single Noisy Point Clouds through Finetuning Data-Driven based PriorsChao Chen, Yu-Shen Liu, Zhizhong HanNeurIPS 2024 · 8 citations
- SGAD-SLAM: Splatting Gaussians at Adjusted Depth for Better Radiance Fields in RGBD SLAMPengchong Hu, Zhizhong HanCVPR 2026 · 2 citations
- Sensing Surface Patches in Volume Rendering for Inferring Signed Distance FunctionsSijia Jiang, Tong Wu, Jing Hua, Zhizhong HanAAAI 2025 · 2 citations
Builds on57
- Instant neural graphics primitives with a multiresolution hash encodingThomas Müller, Alex Evans, Christoph Schied, Alexander KellerSIGGRAPH 2022 · 4,089 citations
- NeuS: Learning Neural Implicit Surfaces by Volume Rendering for Multi-view ReconstructionPeng Wang, Lingjie Liu, Yuan Liu, Christian Theobalt et al.NeurIPS 2021 · 2,500 citations
- Volume Rendering of Neural Implicit SurfacesLior Yariv, Jiatao Gu, Yoni Kasten, Yaron LipmanNeurIPS 2021 · 1,421 citations
- DROID-SLAM: Deep Visual SLAM for Monocular, Stereo, and RGB-D CamerasZachary Teed, Jia DengNeurIPS 2021 · 1,248 citations
- Plenoxels: Radiance Fields without Neural NetworksSara Fridovich-Keil, Alex Yu, Matthew Tancik, Qinhong Chen et al.CVPR 2022 · 1,237 citations
Related papers
- BNV-Fusion: Dense 3D Reconstruction using Bi-level Neural Volume FusionKejie Li, Yansong Tang, Victor Adrian Prisacariu, Philip H. S. TorrCVPR 2022 · 35 citations
- Dense RGB Slam with Neural Implicit MapsHeng Li, Xiaodong Gu, Weihao Yuan, Luwei Yang et al.ICLR 2023 · 10 citations
- SNI-SLAM: Semantic Neural Implicit SLAMSiting Zhu, Guangming Wang, Hermann Blum, Jiuming Liu et al.CVPR 2024 · 59 citations
- SAR-SLAM: Self-Attentive Rendering-based SLAM with Neural Point Cloud EncodingXudong Lv, Zhiwei He, Yuxiang Yang, Jiahao Nie et al.ACM MM 2024 · 2 citations
- Neural RGB-D Surface ReconstructionDejan Azinovic, Ricardo Martin-Brualla, Dan B. Goldman, Matthias Nießner et al.CVPR 2022 · 272 citations
