Graph-Based Prediction and Planning Policy Network (GP3Net) for Scalable Self-Driving in Dynamic Environments Using Deep Reinforcement Learning
Jayabrata Chowdhury, Venkataramanan Shivaraman, Suresh Sundaram, P. B. Sujit
Abstract
Recent advancements in motion planning for Autonomous Vehicles (AVs) show great promise in using expert driver behaviors in non-stationary driving environments. However, learning only through expert drivers needs more generalizability to recover from domain shifts and near-failure scenarios due to the dynamic behavior of traffic participants and weather conditions. A deep Graph-based Prediction and Planning Policy Network (GP3Net) framework is proposed for non-stationary environments that encodes the interactions between traffic participants with contextual information and provides a decision for safe maneuver for AV. A spatio-temporal graph models the interactions between traffic participants for predicting the future trajectories of those participants. The predicted trajectories are utilized to generate a future occupancy map around the AV with uncertainties embedded to anticipate the evolving non-stationary driving environments. Then the contextual information and future occupancy maps are input to the policy network of the GP3Net framework and trained using Proximal Policy Optimization (PPO) algorithm. The proposed GP3Net performance is evaluated on standard CARLA benchmarking scenarios with domain shifts of traffic patterns (urban, highway, and mixed). The results show that the GP3Net outperforms previous state-of-the-art imitation learning-based planning models for different towns. Further, in unseen new weather conditions, GP3Net completes the desired route with fewer traffic infractions. Finally, the results emphasize the advantage of including the prediction module to enhance safety measures in non-stationary environments.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 47b7502f-eee1-4472-8005-f05bbdd05fd5Cited by top-tier papers1
Ask how each one uses itBuilds on7
- Spatial-Temporal Fusion Graph Neural Networks for Traffic Flow ForecastingMengzhang Li, Zhanxing ZhuAAAI 2021 · 1,037 citations
- Exploring the Limitations of Behavior Cloning for Autonomous DrivingFelipe Codevilla, Eder Santana, Antonio M. López, Adrien GaidonICCV 2019 · 666 citations
- The Trajectron: Probabilistic Multi-Agent Trajectory Modeling With Dynamic Spatiotemporal GraphsBoris Ivanovic, Marco PavoneICCV 2019 · 473 citations
- Deep Imitative Models for Flexible Inference, Planning, and ControlNicholas Rhinehart, Rowan McAllister, Sergey LevineICLR 2020 · 159 citations
- Coopernaut: End-to-End Driving with Cooperative Perception for Networked VehiclesJiaxun Cui, Hang Qiu, Dian Chen, Peter Stone et al.CVPR 2022 · 109 citations
Related papers
- HPNet: Dynamic Trajectory Forecasting with Historical Prediction AttentionXiaolong Tang, Meina Kan, Shiguang Shan, Zhilong Ji et al.CVPR 2024 · 66 citations
- Prediction by Anticipation: An Action-Conditional Prediction Method based on Interaction LearningErshad Banijamali, Mohsen Rohani, Elmira Amirloo Abolfathi, Jun Luo et al.ICCV 2021 · 4 citations
- VADv2: End-to-End Vectorized Autonomous Driving via Probabilistic PlanningBo Jiang, Shaoyu Chen, Hao Gao, Bencheng Liao et al.ICLR 2026 · 259 citations
- GoIRL: Graph-Oriented Inverse Reinforcement Learning for Multimodal Trajectory PredictionMuleilan Pei, Shaoshuai Shi, Lu Zhang, Peiliang Li et al.ICML 2025
- OccDriver: Future Occupancy Guided Dual-branch Trajectory Planner in Autonomous DrivingZhao Huang, Bowen Zhang, Zhongzhu Li, Di LinICLR 2026
