Learning Neural Implicit through Volume Rendering with Attentive Depth Fusion Priors
Pengchong Hu, Zhizhong Han
摘要
Learning neural implicit representations has achieved remarkable performance in 3D reconstruction from multi-view images. Current methods use volume rendering to render implicit representations into either RGB or depth images that are supervised by multi-view ground truth. However, rendering a view each time suffers from incomplete depth at holes and unawareness of occluded structures from the depth supervision, which severely affects the accuracy of geometry inference via volume rendering. To resolve this issue, we propose to learn neural implicit representations from multi-view RGBD images through volume rendering with an attentive depth fusion prior. Our prior allows neural networks to perceive coarse 3D structures from the Truncated Signed Distance Function (TSDF) fused from all depth images available for rendering. The TSDF enables accessing the missing depth at holes on one depth image and the occluded parts that are invisible from the current view. By introducing a novel attention mechanism, we allow neural networks to directly use the depth fusion prior with the inferred occupancy as the learned implicit function. Our attention mechanism works with either a one-time fused TSDF that represents a whole scene or an incrementally fused TSDF that represents a partial scene in the context of Simultaneous Localization and Mapping (SLAM). Our evaluations on widely used benchmarks including synthetic and real-world scans show our superiority over the latest neural implicit methods. Please see our project page for code and data at https://machineperceptionlab.github.io/Attentive_DF_Prior/ .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- Neural Signed Distance Function Inference through Splatting 3D Gaussians Pulled on Zero-Level SetWenyuan Zhang, Yu-Shen Liu, Zhizhong HanNeurIPS 2024 · 被引用 58 次
- MultiPull: Detailing Signed Distance Functions by Pulling Multi-Level Queries at Multi-StepTakeshi Noda, Chao Chen, Weiqi Zhang, Xinhai Liu 等NeurIPS 2024 · 被引用 19 次
- Inferring Neural Signed Distance Functions by Overfitting on Single Noisy Point Clouds through Finetuning Data-Driven based PriorsChao Chen, Yu-Shen Liu, Zhizhong HanNeurIPS 2024 · 被引用 8 次
- SGAD-SLAM: Splatting Gaussians at Adjusted Depth for Better Radiance Fields in RGBD SLAMPengchong Hu, Zhizhong HanCVPR 2026 · 被引用 2 次
- Sensing Surface Patches in Volume Rendering for Inferring Signed Distance FunctionsSijia Jiang, Tong Wu, Jing Hua, Zhizhong HanAAAI 2025 · 被引用 2 次
它引用的顶会 Paper57
- Instant neural graphics primitives with a multiresolution hash encodingThomas Müller, Alex Evans, Christoph Schied, Alexander KellerSIGGRAPH 2022 · 被引用 4,089 次
- NeuS: Learning Neural Implicit Surfaces by Volume Rendering for Multi-view ReconstructionPeng Wang, Lingjie Liu, Yuan Liu, Christian Theobalt 等NeurIPS 2021 · 被引用 2,500 次
- Volume Rendering of Neural Implicit SurfacesLior Yariv, Jiatao Gu, Yoni Kasten, Yaron LipmanNeurIPS 2021 · 被引用 1,421 次
- DROID-SLAM: Deep Visual SLAM for Monocular, Stereo, and RGB-D CamerasZachary Teed, Jia DengNeurIPS 2021 · 被引用 1,248 次
- Plenoxels: Radiance Fields without Neural NetworksSara Fridovich-Keil, Alex Yu, Matthew Tancik, Qinhong Chen 等CVPR 2022 · 被引用 1,237 次
相关 Paper
- BNV-Fusion: Dense 3D Reconstruction using Bi-level Neural Volume FusionKejie Li, Yansong Tang, Victor Adrian Prisacariu, Philip H. S. TorrCVPR 2022 · 被引用 35 次
- Dense RGB Slam with Neural Implicit MapsHeng Li, Xiaodong Gu, Weihao Yuan, Luwei Yang 等ICLR 2023 · 被引用 10 次
- SNI-SLAM: Semantic Neural Implicit SLAMSiting Zhu, Guangming Wang, Hermann Blum, Jiuming Liu 等CVPR 2024 · 被引用 59 次
- SAR-SLAM: Self-Attentive Rendering-based SLAM with Neural Point Cloud EncodingXudong Lv, Zhiwei He, Yuxiang Yang, Jiahao Nie 等ACM MM 2024 · 被引用 2 次
- Neural RGB-D Surface ReconstructionDejan Azinovic, Ricardo Martin-Brualla, Dan B. Goldman, Matthias Nießner 等CVPR 2022 · 被引用 272 次
