Localization is All You Evaluate: Data Leakage in Online Mapping Datasets and How to Fix it
Adam Lilja, Junsheng Fu, Erik Stenborg, Lars Hammarstrand
Abstract
The task of online mapping is to predict a local map using current sensor observations, e.g. from lidar and camera, without relying on a pre-built map. State-of-the-art methods are based on supervised learning and are trained predominantly using two datasets: nuScenes and Argoverse 2. However, these datasets revisit the same geographic locations across training, validation, and test sets. Specifically, over 80% of nuScenes and 40% of Argoverse 2 validation and test samples are less than 5 m from a training sample. At test time, the methods are thus evaluated more on how well they localize within a memorized implicit map built from the training data than on extrapolating to unseen locations. Naturally, this data leakage causes inflated performance numbers and we propose geographically disjoint data splits to reveal the true performance in unseen environments. Experimental results show that methods perform considerably worse, some dropping more than 45 mAP, when trained and evaluated on proper data splits. Additionally, a reassessment of prior design choices reveals diverging conclusions from those based on the original split. Notably, the impact of lifting methods and the support from auxiliary tasks (e.g., depth supervision) on performance appears less substantial or follows a different trajectory than previously perceived.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e035f160-b6c9-4ef6-8dc1-9a2e296bbfd4Cited by top-tier papers7
- Stability Under Scrutiny: Benchmarking Representation Paradigms for Online HD MappingHao Shan, Ruikai Li, Han Jiang, Yizhe Fan et al.ICLR 2026 · 10 citations
- SDTagNet: Leveraging Text-Annotated Navigation Maps for Online HD Map ConstructionFabian Immel, Jan-Hendrik Pauls, Richard Schwarzkopf, Frank Bieder et al.NeurIPS 2025 · 9 citations
- PseudoMapTrainer: Learning Online Mapping without HD MapsChristian Löwens, Thorben Funke, Jingchao Xie, Alexandru Paul ConduracheICCV 2025 · 5 citations
- Towards Foundational Models for Single-Chip RadarTianshu Huang, Akarsh Prabhakara, Chuhan Chen, Jay Karhade et al.ICCV 2025 · 3 citations
- ArgoTweak: Towards Self-Updating HD Maps Through Structured PriorsLena Wild, Rafael Valencia, Patric JensfeltICCV 2025 · 2 citations
Builds on13
- RandAugment: Practical Automated Data Augmentation with a Reduced Search SpaceEkin Dogus Cubuk, Barret Zoph, Jonathon Shlens, Quoc LeNeurIPS 2020 · 4,453 citations
- VectorMapNet: End-to-end Vectorized HD Map LearningYicheng Liu, Tianyuan Yuan, Yue Wang, Yilun Wang et al.ICML 2023 · 332 citations
- Cross-view Transformers for real-time Map-view Semantic SegmentationBrady Zhou, Philipp KrähenbühlCVPR 2022 · 279 citations
- PolarFormer: Multi-Camera 3D Object Detection with Polar TransformerYanqin Jiang, Li Zhang, Zhenwei Miao, Xiatian Zhu et al.AAAI 2023 · 240 citations
- Structured Bird's-Eye-View Traffic Scene Understanding from Onboard ImagesYigit Baran Can, Alexander Liniger, Danda Pani Paudel, Luc Van GoolICCV 2021 · 147 citations
Related papers
- Failure Modes for Deep Learning-Based Online Mapping: How to Measure and Address ThemMichael Hubbertz, Qi Han, Tobias MeisenCVPR 2026 · 1 citation
- Producing and Leveraging Online Map Uncertainty in Trajectory PredictionXunjiang Gu, Guanyu Song, Igor Gilitschenski, Marco Pavone et al.CVPR 2024
- Spatial Retrieval Augmented Autonomous DrivingXiaosong Jia, Chenhe Zhang, Yule Jiang, Songbur Wong et al.CVPR 2026 · 6 citations
- AMap: Distilling Future Priors for Ahead-Aware Online HD Map ConstructionRuikai Li, Xinrun Li, Mengwei Xie, Hao Shan et al.CVPR 2026 · 8 citations
- 3DSFLabelling: Boosting 3D Scene Flow Estimation by Pseudo Auto-LabellingChaokang Jiang, Guangming Wang, Jiuming Liu, Hesheng Wang et al.CVPR 2024
