Exploring Data Aggregation in Policy Learning for Vision-Based Urban Autonomous Driving
Aditya Prakash, Aseem Behl, Eshed Ohn-Bar, Kashyap Chitta, Andreas Geiger
摘要
Data aggregation techniques can significantly improve vision-based policy learning within a training environment, e.g., learning to drive in a specific simulation condition. However, as on-policy data is sequentially sampled and added in an iterative manner, the policy can specialize and overfit to the training conditions. For real-world applications, it is useful for the learned policy to generalize to novel scenarios that differ from the training conditions. To improve policy learning while maintaining robustness when training end-to-end driving policies, we perform an extensive analysis of data aggregation techniques in the CARLA environment. We demonstrate how the majority of them have poor generalization performance, and develop a novel approach with empirically better generalization performance compared to existing techniques. Our two key ideas are (1) to sample critical states from the collected on-policy data based on the utility they provide to the learned policy in terms of driving behavior, and (2) to incorporate a replay buffer which progressively focuses on the high uncertainty regions of the policy's state distribution. We evaluate the proposed approach on the CARLA NoCrash benchmark, focusing on the most challenging driving scenarios with dense pedestrian and vehicle traffic. Our approach improves driving success rate by 16% over stateof-the-art, achieving 87% of the expert performance while also reducing the collision rate by an order of magnitude without the use of any additional modality, auxiliary tasks, architectural modifications or reward from the environment.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper11
- Trajectory-guided Control Prediction for End-to-end Autonomous Driving: A Simple yet Strong BaselinePenghao Wu, Xiaosong Jia, Li Chen, Junchi Yan 等NeurIPS 2022 · 被引用 444 次
- End-to-End Urban Driving by Imitating a Reinforcement Learning CoachZhejun Zhang, Alexander Liniger, Dengxin Dai, Fisher Yu 等ICCV 2021 · 被引用 313 次
- NEAT: Neural Attention Fields for End-to-End Autonomous DrivingKashyap Chitta, Aditya Prakash, Andreas GeigerICCV 2021 · 被引用 274 次
- DriveAdapter: Breaking the Coupling Barrier of Perception and Planning in End-to-End Autonomous DrivingXiaosong Jia, Yulu Gao, Li Chen, Junchi Yan 等ICCV 2023 · 被引用 154 次
- Hidden Biases of End-to-End Driving ModelsBernhard Jaeger, Kashyap Chitta, Andreas GeigerICCV 2023 · 被引用 130 次
它引用的顶会 Paper1
相关 Paper
- EE-RL: Vision Language Guided Reinforcement Learning with Explorer and Expert model for End-to-End Autonomous DrivingXiaolong Li, Lan Yang, Ruyang Li, Shan Fang 等CVPR 2026
- Learning to drive from a world on railsDian Chen, Vladlen Koltun, Philipp KrähenbühlICCV 2021 · 被引用 164 次
- Learning from All VehiclesDian Chen, Philipp KrähenbühlCVPR 2022
- LEAD: Minimizing Learner-Expert Asymmetry in End-to-End DrivingLong Nguyen, Micha Fauth, Bernhard Jaeger, Daniel Dauner 等CVPR 2026 · 被引用 28 次
- Learning by WatchingJimuyang Zhang, Eshed Ohn-BarCVPR 2021
