MHED-SLAM: Multi-Scale Hybrid Encoding-Based Decoupled SLAM
Dengfang Feng, Wenyang Qin, Zhongchen Shi, Wei Chen, Yanhui Duan, Liang Xie, Erwei Yin
摘要
Neural Radiance Fields (NeRF)-based Visual Simultaneous Localization and Mapping (SLAM) achieve superior scene geometric modeling and robust camera tracking by leveraging neural representations. Existing methods typically relied on multi-resolution hash encoding with truncated signed distance fields (TSDF) to achieve high frame rates. However, unavoidable hash collisions can lead to artifacts, and multi-view color inconsistencies in indoor scenes can result in shape-radiance ambiguity, adversely affecting geometric quality and tracking accuracy. To address these issues, we propose a novel Multi-scale Hybrid Encoding-based Decoupled SLAM (MHED-SLAM). First, to mitigate the adverse effects of hash collisions and reduce the number of learnable parameters, we innovatively fuse a coarse-scale hash tri-plane with a fine-scale hash grid within a single latent volume. Second, to enable precise geometric reconstruction and camera tracking, we decouple the reconstruction and rendering processes, independently learning a TSDF field for reconstruction and a density field for rendering. Third, we devise a Symmetric Kullback-Leibler (SKL) strategy based on ray termination distributions to align the probability distributions derived from the TSDF and density fields for their synchronous convergence. Extensive experimental evaluations demonstrate that our approach surpasses the state-of-the-art (SOTA) methods by utilizing a faster frame rate of 20 Hz and fewer parameters, while achieving higher tracking and reconstruction accuracy.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper19
- Instant neural graphics primitives with a multiresolution hash encodingThomas Müller, Alex Evans, Christoph Schied, Alexander KellerSIGGRAPH 2022 · 被引用 4,089 次
- NeuS: Learning Neural Implicit Surfaces by Volume Rendering for Multi-view ReconstructionPeng Wang, Lingjie Liu, Yuan Liu, Christian Theobalt 等NeurIPS 2021 · 被引用 2,500 次
- Volume Rendering of Neural Implicit SurfacesLior Yariv, Jiatao Gu, Yoni Kasten, Yaron LipmanNeurIPS 2021 · 被引用 1,421 次
- iMAP: Implicit Mapping and Positioning in Real-TimeEdgar Sucar, Shikun Liu, Joseph Ortiz, Andrew J. DavisonICCV 2021 · 被引用 834 次
- Depth-supervised NeRF: Fewer Views and Faster Training for FreeKangle Deng, Andrew Liu, Jun-Yan Zhu, Deva RamananCVPR 2022 · 被引用 756 次
相关 Paper
- ESLAM: Efficient Dense SLAM System Based on Hybrid Representation of Signed Distance FieldsMohammad Mahdi Johari, Camilla Carta, François FleuretCVPR 2023
- SAR-SLAM: Self-Attentive Rendering-based SLAM with Neural Point Cloud EncodingXudong Lv, Zhiwei He, Yuxiang Yang, Jiahao Nie 等ACM MM 2024 · 被引用 2 次
- 3D Reconstruction and Novel View Synthesis of Indoor Environments Based on a Dual Neural Radiance FieldZhenyu Bao, Guibiao Liao, Zhongyuan Zhao, Kanglin Liu 等ACM MM 2024 · 被引用 3 次
- IBD-SLAM: Learning Image-Based Depth Fusion for Generalizable SLAMMinghao Yin, Shangzhe Wu, Kai HanCVPR 2024 · 被引用 5 次
- Robust Camera Pose Refinement for Multi-Resolution Hash EncodingHwan Heo, Taekyung Kim, Jiyoung Lee, Jaewon Lee 等ICML 2023 · 被引用 26 次
