MP3: A Unified Model To Map, Perceive, Predict and Plan
Sergio Casas, Abbas Sadat, Raquel Urtasun
Abstract
High-definition maps (HD maps) are a key component of most modern self-driving systems due to their valuable semantic and geometric information. Unfortunately, building HD maps has proven hard to scale due to their cost as well as the requirements they impose in the localization system that has to work everywhere with centimeter-level accuracy. Being able to drive without an HD map would be very beneficial to scale self-driving solutions as well as to increase the failure tolerance of existing ones (e.g., if localization fails or the map is not up-to-date). Towards this goal, we propose MP3, an end-to-end approach to mapless 1 driving where the input is raw sensor data and a high-level command (e.g., turn left at the intersection). MP3 predicts intermediate representations in the form of an online map and the current and future state of dynamic agents, and exploits them in a novel neural motion planner to make interpretable decisions taking into account uncertainty. We show that our approach is significantly safer, more comfortable, and can follow commands better than the baselines in challenging long-term closed-loop simulations, as well as when compared to an expert driver in a large-scale real-world dataset.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ee6b1d7e-9b35-4d52-81f5-8369668eb8b6Cited by top-tier papers70
- VAD: Vectorized Scene Representation for Efficient Autonomous DrivingBo Jiang, Shaoyu Chen, Qing Xu, Bencheng Liao et al.ICCV 2023 · 602 citations
- Motion Transformer with Global Intention Localization and Local Movement RefinementShaoshuai Shi, Li Jiang, Dengxin Dai, Bernt SchieleNeurIPS 2022 · 515 citations
- Trajectory-guided Control Prediction for End-to-end Autonomous Driving: A Simple yet Strong BaselinePenghao Wu, Xiaosong Jia, Li Chen, Junchi Yan et al.NeurIPS 2022 · 444 citations
- Vista: A Generalizable Driving World Model with High Fidelity and Versatile ControllabilityShenyuan Gao, Jiazhi Yang, Li Chen, Kashyap Chitta et al.NeurIPS 2024 · 403 citations
- VectorMapNet: End-to-end Vectorized HD Map LearningYicheng Liu, Tianyuan Yuan, Yue Wang, Yilun Wang et al.ICML 2023 · 332 citations
Builds on8
- Exploring the Limitations of Behavior Cloning for Autonomous DrivingFelipe Codevilla, Eder Santana, Antonio M. López, Adrien GaidonICCV 2019 · 666 citations
- 3D-LaneNet: End-to-End 3D Multiple Lane DetectionNoa Garnett, Rafi Cohen, Tomer Pe'er, Roee Lahav et al.ICCV 2019 · 232 citations
- Deep Imitative Models for Flexible Inference, Planning, and ControlNicholas Rhinehart, Rowan McAllister, Sergey LevineICLR 2020 · 159 citations
- MotionNet: Joint Perception and Motion Prediction for Autonomous Driving Based on Bird's Eye View MapsPengxiang Wu, Siheng Chen, Dimitris N. MetaxasCVPR 2020
- CoverNet: Multimodal Behavior Prediction Using Trajectory SetsTung Phan-Minh, Elena Corina Grigore, Freddy A. Boulton, Oscar Beijbom et al.CVPR 2020
Related papers
- Self-Supervised Simultaneous Multi-Step Prediction of Road Dynamics and Cost MapElmira Amirloo Abolfathi, Mohsen Rohani, Ershad Banijamali, Jun Luo et al.CVPR 2021
- RTMap: Real-Time Recursive Mapping with Change Detection and LocalizationYuheng Du, Sheng Yang, Lingxuan Wang, Zhenghua Hou et al.ICCV 2025 · 2 citations
- MapTR: Structured Modeling and Learning for Online Vectorized HD Map ConstructionBencheng Liao, Shaoyu Chen, Xinggang Wang, Tianheng Cheng et al.ICLR 2023 · 69 citations
- Scene Reconstruction as Mapping Priors for 3D DetectionYang Fu, Yuliang Zou, Hao Xiang, Xin Huang et al.CVPR 2026 · 1 citation
- HDMapGen: A Hierarchical Graph Generative Model of High Definition MapsLu Mi, Hang Zhao, Charlie Nash, Xiaohan Jin et al.CVPR 2021
