U-ViLAR: Uncertainty-Aware Visual Localization for Autonomous Driving via Differentiable Association and Registration
Xiaofan Li, Zhihao Xu, Chenming Wu, Zhao Yang, Yumeng Zhang, Jiang-Jiang Liu, Haibao Yu, Xiaoqing Ye, Yuan Wang, Shirui Li, Xun Sun, Ji Wan, Jun Wang
Abstract
Accurate localization using visual information is a critical yet challenging task, especially in urban environments where nearby buildings and construction sites significantly degrade GNSS (Global Navigation Satellite System) signal quality. This issue underscores the importance of visual localization techniques in scenarios where GNSS signals are unreliable. This paper proposes U-ViLAR, a novel uncertainty-aware visual localization framework designed to address these challenges while enabling adaptive localization using high-definition (HD) maps or navigation maps. Specifically, our method first extracts features from the input visual data and maps them into Bird's-Eye-View (BEV) space to enhance spatial consistency with the map input. Subsequently, we introduce: a) Perceptual Uncertainty-guided Association, which mitigates errors caused by perception uncertainty, and b) Localization Uncertainty-guided Registration, which reduces errors introduced by localization uncertainty. By effectively balancing the coarse-grained large-scale localization capability of association with the fine-grained precise localization capability of registration, our approach achieves robust and accurate localization. Experimental results demonstrate that our method achieves state-of-the-art performance across multiple localization tasks. Furthermore, our model has undergone rigorous testing on large-scale autonomous driving fleets and has demonstrated stable performance in various challenging urban scenarios.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ff1c6e50-6500-483d-9f8c-6a95a8ef5508Cited by top-tier papers2
- FVAR: Next-Focus Prediction for Visual Autoregressive ModelingXiaofan Li, Chenming Wu, Yanpeng Sun, Jiaming Zhou et al.CVPR 2026
- DriVerse: Navigation World Model for Driving Simulation via Multimodal Trajectory Prompting and Motion AlignmentXiaofan Li, Chenming Wu, Zhao Yang, Zhihao Xu et al.ACM MM 2025
Builds on13
- BEVDepth: Acquisition of Reliable Depth for Multi-View 3D Object DetectionYinhao Li, Zheng Ge, Guanyi Yu, Jinrong Yang et al.AAAI 2023 · 954 citations
- VectorMapNet: End-to-end Vectorized HD Map LearningYicheng Liu, Tianyuan Yuan, Yue Wang, Yilun Wang et al.ICML 2023 · 332 citations
- Pixel-Perfect Structure-from-Motion with Featuremetric RefinementPhilipp Lindenberger, Paul-Edouard Sarlin, Viktor Larsson, Marc PollefeysICCV 2021 · 266 citations
- Beyond Cross-view Image Retrieval: Highly Accurate Vehicle Localization Using Satellite ImageYujiao Shi, Hongdong LiCVPR 2022 · 81 citations
- VRP-SAM: SAM with Visual Reference PromptYanpeng Sun, Jiahui Chen, Shan Zhang, Xinyu Zhang et al.CVPR 2024 · 49 citations
Related papers
- VILAM: Infrastructure-assisted 3D Visual Localization and Mapping for Autonomous DrivingJiahe Cui, Shuyao Shi, Yuze He, Jianwei Niu et al.NSDI 2024 · 17 citations
- HOLO: Homography-Guided Pose Estimator Network for Fine-Grained Visual Localization on SD MapsXuchang Zhong, Xu Cao, Jinke Feng, Hao FangCVPR 2026
- Efficient Large-scale Localization by Global Instance RecognitionFei Xue, Ignas Budvytis, Daniel Olmeda Reino, Roberto CipollaCVPR 2022 · 19 citations
- SFD2: Semantic-Guided Feature Detection and DescriptionFei Xue, Ignas Budvytis, Roberto CipollaCVPR 2023
- MapUQ: Map with Uncertainty Quantification for Robust BEV Vectorized ConstructionShaoyuan Mo, MaQi, l r, Bohan Li et al.ICML 2026
