Model-Based Policy Adaptation for Closed-Loop End-to-end Autonomous Driving
Haohong Lin, Yunzhi Zhang, Wenhao Ding, Jiajun Wu, Ding Zhao
摘要
End-to-end (E2E) autonomous driving models have demonstrated strong performance in open-loop evaluations but often suffer from cascading errors and poor generalization in closed-loop settings. To address this gap, we propose Model-based Policy Adaptation (MPA), a general framework that enhances the robustness and safety of pretrained E2E driving agents during deployment. MPA first generates diverse counterfactual trajectories using a geometry-consistent simulation engine, exposing the agent to scenarios beyond the original dataset. Based on this generated data, MPA trains a diffusion-based policy adapter to refine the base policy's predictions and a multi-step Q value model to evaluate long-term outcomes. At inference time, the adapter proposes multiple trajectory candidates, and the Q value model selects the one with the highest expected utility. Experiments on the nuScenes benchmark using a photorealistic closed-loop simulator demonstrate that MPA significantly improves performance across in-domain, out-of-domain, and safety-critical scenarios. We further investigate how the scale of counterfactual data and inference-time guidance strategies affect overall effectiveness.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- PlannerRFT: Reinforcing Diffusion Planners through Closed-Loop and Sample-Efficient Fine-TuningHongchen Li, Tianyu Li, Jiazhi Yang, Mingyang Shang 等CVPR 2026 · 被引用 13 次
- Drive My Way: Preference Alignment of Vision-Language-Action Model for Personalized DrivingZehao Wang, Huaide Jiang, Shuaiwu Dong, Yuping Wang 等CVPR 2026 · 被引用 7 次
它引用的顶会 Paper23
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 被引用 11,743 次
- 3D Gaussian Splatting for Real-Time Radiance Field RenderingBernhard Kerbl, Georgios Kopanas, Thomas Leimkühler, George DrettakisSIGGRAPH 2023 · 被引用 5,687 次
- Planning with Diffusion for Flexible Behavior SynthesisMichael Janner, Yilun Du, Joshua B. Tenenbaum, Sergey LevineICML 2022 · 被引用 1,115 次
- VAD: Vectorized Scene Representation for Efficient Autonomous DrivingBo Jiang, Shaoyu Chen, Qing Xu, Bencheng Liao 等ICCV 2023 · 被引用 602 次
- Vista: A Generalizable Driving World Model with High Fidelity and Versatile ControllabilityShenyuan Gao, Jiazhi Yang, Li Chen, Kashyap Chitta 等NeurIPS 2024 · 被引用 403 次
相关 Paper
- Driving with Advice: Large Model as Motion Advisor for Joint PlanningJunyin Wang, Jinlei Yu, Hao Lin, Huikai Liu 等AAAI 2026
- Driving View Synthesis on Free-Form Trajectories with Generative PriorZeyu Yang, Zijie Pan, Yuankun Yang, Xiatian Zhu 等ICCV 2025 · 被引用 3 次
- DiffusionDrive: Truncated Diffusion Model for End-to-End Autonomous DrivingBencheng Liao, Shaoyu Chen, Haoran Yin, Bo Jiang 等CVPR 2025
- Diffusion-Based Planning for Autonomous Driving with Flexible GuidanceYinan Zheng, Ruiming Liang, Kexin Zheng, Jinliang Zheng 等ICLR 2025
- Unraveling the Effects of Synthetic Data on End-to-End Autonomous DrivingJunhao Ge, Zuhong Liu, Longteng Fan, Yifan Jiang 等ICCV 2025 · 被引用 2 次
