VILAM: Infrastructure-assisted 3D Visual Localization and Mapping for Autonomous Driving
Jiahe Cui, Shuyao Shi, Yuze He, Jianwei Niu, Guoliang Xing, Zhenchao Ouyang
Abstract
Visual Simultaneous Localization and Mapping (SLAM) presents a promising avenue for fulfilling the essential perception and localization tasks in autonomous driving systems using cost-effective visual sensors. Nevertheless, existing visual SLAM frameworks often suffer from substantial cumulative errors and performance degradation in complicated driving scenarios. In this paper, we propose VILAM, a novel framework that leverages intelligent roadside infrastructures to realize high-precision and globally consistent localization and mapping on autonomous vehicles. The key idea of VILAM is to utilize the precise scene measurement from the infrastructure as global references to correct errors in the local map constructed by the vehicle. To overcome the unique deformation in the 3D local map to align it with the infrastructure measurement, VILAM proposes a novel elastic point cloud registration method that enables independent optimization of different parts of the local map. Moreover, VILAM adopts a lightweight factor graph construction and optimization to first correct the vehicle trajectory, and thus reconstruct the consistent global map efficiently. We implement the VILAM end-to-end on a real-world smart lamppost testbed in multiple road scenarios. Extensive experiment results show that VILAM can achieve decimeter-level localization and mapping accuracy with consumer-level onboard cameras and is robust under diverse road scenarios. A video demo of VILAM on our real-world testbed is available at https://youtu.be/lTlqDNipDVE.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d1fc6ac1-42ae-4f93-93db-53ed145c3449Cited by top-tier papers2
- 3D Bounding Box Estimation Based on COTS mmWave Radar via Moving ScanningYiwen Feng, Jiayang Zhao, Chuyu Wang, Lei Xie et al.UbiComp 2025 · 8 citations
- Towards Real-Time Defense against Object-Based LiDAR Attacks in Autonomous DrivingYan Zhang, Zihao Liu, Yi Zhu, Chenglin MiaoCCS 2025
Builds on7
- NICE-SLAM: Neural Implicit Scalable Encoding for SLAMZihan Zhu, Songyou Peng, Viktor Larsson, Weiwei Xu et al.CVPR 2022 · 720 citations
- VIPS: real-time perception fusion for infrastructure-assisted autonomous drivingShuyao Shi, Jiahe Cui, Zhehao Jiang, Zhenyu Yan et al.MobiCom 2022 · 126 citations
- CarMap: Fast 3D Feature Map Updates for AutomobilesFawad Ahmad, Hang Qiu, Ray Eells, Fan Bai et al.NSDI 2020 · 87 citations
- VI-eye: semantic-based 3D point cloud registration for infrastructure-assisted autonomous drivingYuze He, Li Ma, Zhehao Jiang, Yi Tang et al.MobiCom 2021 · 76 citations
- SwarmMap: Scaling Up Real-time Collaborative Visual SLAM at the EdgeJingao Xu, Hao Cao, Zheng Yang, Longfei Shangguan et al.NSDI 2022 · 71 citations
Related papers
- VI-Map: Infrastructure-Assisted Real-Time HD Mapping for Autonomous DrivingYuze He, Chen Bian, Jingfei Xia, Shuyao Shi et al.MobiCom 2023 · 35 citations
- U-ViLAR: Uncertainty-Aware Visual Localization for Autonomous Driving via Differentiable Association and RegistrationXiaofan Li, Zhihao Xu, Chenming Wu, Zhao Yang et al.ICCV 2025
- MamV2XCalib: V2X-based Target-Less Infrastructure Camera Calibration with State Space ModelYaoye Zhu, Zhe Wang, Yan WangICCV 2025 · 3 citations
- Adversary is on the Road: Attacks on Visual SLAM using Unnoticeable Adversarial PatchBaodong Chen, Wei Wang, Pascal Sikorski, Ting ZhuUSENIX Security 2024 · 8 citations
- ESLAM: Efficient Dense SLAM System Based on Hybrid Representation of Signed Distance FieldsMohammad Mahdi Johari, Camilla Carta, François FleuretCVPR 2023
