ColSLAM: A Versatile Collaborative SLAM System for Mobile Phones Using Point-Line Features and Map Caching
Wanting Li, Yongcai Wang, Yongyu Guo, Shuo Wang, Yu Shao, Xuewei Bai, Xudong Cai, Qiang Ye, Deying Li
Abstract
Over the past years, augmented reality (AR) based on mobile phones has gained great attention. When multiple phones are used in AR applications, collaborative simultaneous localization and mapping (SLAM) is considered one of the enabling technologies, i.e., multiple mobile phones complete the localization and mapping through collaboration. However, the state-of-the-art collaborative SLAM systems not only suffer from the delays introduced by a high-complexity graph optimization problem, but also may exhibit varying levels of accuracy across dissimilar environments or different types of mobile devices. In this paper, we propose a scalable and robust collaborative SLAM system, point-line-based Collaborative SLAM (ColSLAM). Technically, ColSLAM includes two innovative features that help achieve satisfactory scalability and robustness. First, a mapping cacher (MC) is designed for each agent on the server, which uses global keyframes to detect loop closures, updates the cached local map, and quickly responds to the agent's pose drifts. With MC, each agent's local pose is corrected using global knowledge in real-time. Secondly, to improve the robustness performance, ColSLAM employs point-line-fusion-based Visual Inertial Odometry (VIO), point-line-fusion-based NetVLAD loop detection, and an enhanced geometric verification and relative pose calculation method called PNPL. Empirical evaluations based on the EuRoc dataset and real degenerate environments demonstrate that ColSLAM outperforms the existing collaborative SLAM systems in terms of accuracy, robustness, and scalability.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b8cb6ebc-6204-4ddb-a37d-4294115a047bCited by top-tier papers2
- Progress-Think: Semantic Progress Reasoning for Vision-Language NavigationShuo Wang, Yucheng Wang, Guoxin Lian, Yongcai Wang et al.CVPR 2026 · 10 citations
- MambaVO: Deep Visual Odometry Based on Sequential Matching Refinement and Training SmoothingShuo Wang, Wanting Li, Yongcai Wang, Zhaoxin Fan et al.CVPR 2025
Builds on1
Related papers
- 100-Phones: A Large VI-SLAM Dataset for Augmented Reality Towards Mass Deployment on Mobile PhonesGuofeng Zhang, Jin Yuan, Haomin Liu, Zhen Peng et al.IEEE VR 2024 · 5 citations
- Robust Tightly-Coupled Visual-Inertial Odometry with Pre-built Maps in High Latency SituationsHujun Bao, Weijian Xie, Quanhao Qian, Danpeng Chen et al.IEEE VR 2022 · 16 citations
- CP-SLAM: Collaborative Neural Point-based SLAM SystemJiarui Hu, Mao Mao, Hujun Bao, Guofeng Zhang et al.NeurIPS 2023 · 65 citations
- SlimSLAM: An Adaptive Runtime for Visual-Inertial Simultaneous Localization and MappingArmand Behroozi, Yuxiang Chen, Vlad Fruchter, Lavanya Subramanian et al.ASPLOS 2024 · 6 citations
- GSLAMOT: A Tracklet and Query Graph-based Simultaneous Locating, Mapping, and Multiple Object Tracking SystemShuo Wang, Yongcai Wang, Zhimin Xu, Yongyu Guo et al.ACM MM 2024 · 6 citations
