FoundationSLAM: Unleashing the Power of Depth Foundation Models for End-to-End Dense Visual SLAM
Yuchen Wu, Jiahe Li, Fabio Tosi, Matteo Poggi, Jin Zheng, Xiao Bai
摘要
We present FoundationSLAM, a learning-based monocular dense SLAM system that addresses the absence of geometric consistency in previous flow-based approaches for accurate and robust tracking and mapping. Our core idea is to bridge flow estimation with geometric reasoning by leveraging the guidance from foundation depth models. To this end, we first develop a Hybrid Flow Network that produces geometry-aware correspondences, enabling consistent depth and pose inference across diverse keyframes. To enforce global consistency, we propose a Bi-Consistent Bundle Adjustment Layer that jointly optimizes keyframe pose and depth under multi-view constraints. Furthermore, we introduce a Reliability-Aware Refinement mechanism that dynamically adapts the flow update process by distinguishing between reliable and uncertain regions, forming a closed feedback loop between matching and optimization. Extensive experiments demonstrate that FoundationSLAM achieves superior trajectory accuracy and dense reconstruction quality across multiple challenging datasets, while running in real-time at 18 FPS, demonstrating strong generalization to various scenarios and practical applicability of our method.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper15
- Depth Anything V2Lihe Yang, Bingyi Kang, Zilong Huang, Zhen Zhao 等NeurIPS 2024 · 被引用 2,305 次
- DROID-SLAM: Deep Visual SLAM for Monocular, Stereo, and RGB-D CamerasZachary Teed, Jia DengNeurIPS 2021 · 被引用 1,248 次
- NICE-SLAM: Neural Implicit Scalable Encoding for SLAMZihan Zhu, Songyou Peng, Viktor Larsson, Weiwei Xu 等CVPR 2022 · 被引用 720 次
- Gaussian Splatting SLAMHidenobu Matsuki, Riku Murai, Paul H. J. Kelly, Andrew J. DavisonCVPR 2024 · 被引用 328 次
- Deep Patch Visual OdometryZachary Teed, Lahav Lipson, Jia DengNeurIPS 2023 · 被引用 323 次
相关 Paper
- MASt3R-SLAM: Real-Time Dense SLAM with 3D Reconstruction PriorsRiku Murai, Eric Dexheimer, Andrew J. DavisonCVPR 2025
- GO-SLAM: Global Optimization for Consistent 3D Instant ReconstructionYoumin Zhang, Fabio Tosi, Stefano Mattoccia, Matteo PoggiICCV 2023 · 被引用 208 次
- SCE-SLAM: Scale-Consistent Monocular SLAM via Scene Coordinate EmbeddingsYuchen Wu, Jiahe Li, Xiaohan Yu, Lina Yu 等CVPR 2026 · 被引用 1 次
- VGGT-Motion: Motion-Aware Calibration-Free Monocular SLAM for Long-Range ConsistencyZhuang Xiong, Chen Zhang, Qingshan Xu, Wenbing TaoICML 2026 · 被引用 6 次
- Deep Two-View Structure-From-Motion RevisitedJianyuan Wang, Yiran Zhong, Yuchao Dai, Stan Birchfield 等CVPR 2021
