Graph-Based Prediction and Planning Policy Network (GP3Net) for Scalable Self-Driving in Dynamic Environments Using Deep Reinforcement Learning
Jayabrata Chowdhury, Venkataramanan Shivaraman, Suresh Sundaram, P. B. Sujit
摘要
Recent advancements in motion planning for Autonomous Vehicles (AVs) show great promise in using expert driver behaviors in non-stationary driving environments. However, learning only through expert drivers needs more generalizability to recover from domain shifts and near-failure scenarios due to the dynamic behavior of traffic participants and weather conditions. A deep Graph-based Prediction and Planning Policy Network (GP3Net) framework is proposed for non-stationary environments that encodes the interactions between traffic participants with contextual information and provides a decision for safe maneuver for AV. A spatio-temporal graph models the interactions between traffic participants for predicting the future trajectories of those participants. The predicted trajectories are utilized to generate a future occupancy map around the AV with uncertainties embedded to anticipate the evolving non-stationary driving environments. Then the contextual information and future occupancy maps are input to the policy network of the GP3Net framework and trained using Proximal Policy Optimization (PPO) algorithm. The proposed GP3Net performance is evaluated on standard CARLA benchmarking scenarios with domain shifts of traffic patterns (urban, highway, and mixed). The results show that the GP3Net outperforms previous state-of-the-art imitation learning-based planning models for different towns. Further, in unseen new weather conditions, GP3Net completes the desired route with fewer traffic infractions. Finally, the results emphasize the advantage of including the prediction module to enhance safety measures in non-stationary environments.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper7
- Spatial-Temporal Fusion Graph Neural Networks for Traffic Flow ForecastingMengzhang Li, Zhanxing ZhuAAAI 2021 · 被引用 1,037 次
- Exploring the Limitations of Behavior Cloning for Autonomous DrivingFelipe Codevilla, Eder Santana, Antonio M. López, Adrien GaidonICCV 2019 · 被引用 666 次
- The Trajectron: Probabilistic Multi-Agent Trajectory Modeling With Dynamic Spatiotemporal GraphsBoris Ivanovic, Marco PavoneICCV 2019 · 被引用 473 次
- Deep Imitative Models for Flexible Inference, Planning, and ControlNicholas Rhinehart, Rowan McAllister, Sergey LevineICLR 2020 · 被引用 159 次
- Coopernaut: End-to-End Driving with Cooperative Perception for Networked VehiclesJiaxun Cui, Hang Qiu, Dian Chen, Peter Stone 等CVPR 2022 · 被引用 109 次
相关 Paper
- HPNet: Dynamic Trajectory Forecasting with Historical Prediction AttentionXiaolong Tang, Meina Kan, Shiguang Shan, Zhilong Ji 等CVPR 2024 · 被引用 66 次
- Prediction by Anticipation: An Action-Conditional Prediction Method based on Interaction LearningErshad Banijamali, Mohsen Rohani, Elmira Amirloo Abolfathi, Jun Luo 等ICCV 2021 · 被引用 4 次
- VADv2: End-to-End Vectorized Autonomous Driving via Probabilistic PlanningBo Jiang, Shaoyu Chen, Hao Gao, Bencheng Liao 等ICLR 2026 · 被引用 259 次
- GoIRL: Graph-Oriented Inverse Reinforcement Learning for Multimodal Trajectory PredictionMuleilan Pei, Shaoshuai Shi, Lu Zhang, Peiliang Li 等ICML 2025
- OccDriver: Future Occupancy Guided Dual-branch Trajectory Planner in Autonomous DrivingZhao Huang, Bowen Zhang, Zhongzhu Li, Di LinICLR 2026
