Deep Multi-Task Learning for Joint Localization, Perception, and Prediction
John Phillips, Julieta Martinez, Ioan Andrei Barsan, Sergio Casas, Abbas Sadat, Raquel Urtasun
Abstract
Over the last few years, we have witnessed tremendous progress on many subtasks of autonomous driving including perception, motion forecasting, and motion planning. However, these systems often assume that the car is accurately localized against a high-definition map. In this paper we question this assumption, and investigate the issues that arise in state-of-the-art autonomy stacks under localization error. Based on our observations, we design a system that jointly performs perception, prediction, and localization. Our architecture is able to reuse computation between the three tasks, and is thus able to correct localization errors efficiently. We show experiments on a large-scale autonomy dataset, demonstrating the efficiency and accuracy of our proposed approach.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 1ddbbac7-49c2-4bb0-b43e-515964d287eaCited by top-tier papers5
- DeMT: Deformable Mixer Transformer for Multi-Task Learning of Dense PredictionYangyang Xu, Yibo Yang, Lefei ZhangAAAI 2023 · 81 citations
- Does Physical Adversarial Example Really Matter to Autonomous Driving? Towards System-Level Effect of Adversarial Object Evasion AttackNingfei Wang, Yunpeng Luo, Takami Sato, Kaidi Xu et al.ICCV 2023 · 65 citations
- Self-Supervised Bird's Eye View Motion Prediction with Cross-Modality SignalsShaoheng Fang, Zuhong Liu, Mingyu Wang, Chenxin Xu et al.AAAI 2024 · 8 citations
- ViP3D: End-to-End Visual Trajectory Prediction via 3D Agent QueriesJunru Gu, Chenxu Hu, Tianyuan Zhang, Xuanyao Chen et al.CVPR 2023
- Standing Between Past and Future: Spatio-Temporal Modeling for Multi-Camera 3D Multi-Object TrackingZiqi Pang, Jie Li, Pavel Tokmakov, Dian Chen et al.CVPR 2023
Builds on8
- DeepVCP: An End-to-End Deep Neural Network for Point Cloud RegistrationWeixin Lu, Guowei Wan, Yao Zhou, Xiangyu Fu et al.ICCV 2019 · 313 citations
- Learning to Evaluate Perception Models Using Planner-Centric MetricsJonah Philion, Amlan Kar, Sanja FidlerCVPR 2020
- nuScenes: A Multimodal Dataset for Autonomous DrivingHolger Caesar, Varun Bankiti, Alex H. Lang, Sourabh Vora et al.CVPR 2020
- SuperGlue: Learning Feature Matching With Graph Neural NetworksPaul-Edouard Sarlin, Daniel DeTone, Tomasz Malisiewicz, Andrew RabinovichCVPR 2020
- VectorNet: Encoding HD Maps and Agent Dynamics From Vectorized RepresentationJiyang Gao, Chen Sun, Hang Zhao, Yi Shen et al.CVPR 2020
Related papers
- Planning-oriented Autonomous DrivingYihan Hu, Jiazhi Yang, Li Chen, Keyu Li et al.CVPR 2023
- PnPNet: End-to-End Perception and Prediction With Tracking in the LoopMing Liang, Bin Yang, Wenyuan Zeng, Yun Chen et al.CVPR 2020
- Producing and Leveraging Online Map Uncertainty in Trajectory PredictionXunjiang Gu, Guanyu Song, Igor Gilitschenski, Marco Pavone et al.CVPR 2024
- ForeSight: Multi-View Streaming Joint Object Detection and Trajectory ForecastingSandro Papais, Letian Wang, Brian Cheong, Steven L. WaslanderICCV 2025 · 1 citation
- Gaussian YOLOv3: An Accurate and Fast Object Detector Using Localization Uncertainty for Autonomous DrivingJiwoong Choi, Dayoung Chun, Hyun Kim, Hyuk-Jae LeeICCV 2019 · 445 citations
