OAM: An Option-Action Reinforcement Learning Framework for Universal Multi-Intersection Control
Enming Liang, Zicheng Su, Chilin Fang, Renxin Zhong
摘要
Efficient traffic signal control is an important means to alleviate urban traffic congestion. Reinforcement learning (RL) has shown great potentials in devising optimal signal plans that can adapt to dynamic traffic congestion. However, several challenges still need to be overcome. Firstly, a paradigm of state, action, and reward design is needed, especially for an optimality-guaranteed reward function. Secondly, the generalization of the RL algorithms is hindered by the varied topologies and physical properties of intersections. Lastly, enhancing the cooperation between intersections is needed for large network applications. To address these issues, the Option-Action RL framework for universal Multi-intersection control (OAM) is proposed. Based on the well-known cell transmission model, we first define a lane-cell-level state to better model the traffic flow propagation. Based on this physical queuing dynamics, we propose a regularized delay as the reward to facilitate temporal credit assignment while maintaining the equivalence with minimizing the average travel time. We then recapitulate the phase actions as the constrained combinations of lane options and design a universal neural network structure to realize model generalization to any intersection with any phase definition. The multiple-intersection cooperation is then rigorously discussed using the potential game theory.
We test the OAM algorithm under four networks with different settings, including a city-level scenario with 2,048 intersections using synthetic and real-world datasets. The results show that the OAM can outperform the state-of-the-art controllers in reducing the average travel time.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- CoLLMLight: Cooperative Large Language Model Agents for Network-Wide Traffic Signal ControlZirui Yuan, Siqi Lai, Hao LiuICLR 2026 · 被引用 18 次
- AlphaRoute: Large-Scale Coordinated Route Planning via Monte Carlo Tree SearchGuiyang Luo, Yantao Wang, Hui Zhang, Quan Yuan 等AAAI 2023 · 被引用 11 次
它引用的顶会 Paper4
- Weighted QMIX: Expanding Monotonic Value Function Factorisation for Deep Multi-Agent Reinforcement LearningTabish Rashid, Gregory Farquhar, Bei Peng, Shimon WhitesonNeurIPS 2020 · 被引用 1,960 次
- Toward A Thousand Lights: Decentralized Deep Reinforcement Learning for Large-Scale Traffic Signal ControlChacha Chen, Hua Wei, Nan Xu, Guanjie Zheng 等AAAI 2020 · 被引用 450 次
- AttendLight: Universal Attention-Based Reinforcement Learning Model for Traffic Signal ControlAfshin Oroojlooy, MohammadReza Nazari, Davood Hajinezhad, Jorge SilvaNeurIPS 2020 · 被引用 143 次
- Hierarchically and Cooperatively Learning Traffic Signal ControlBingyu Xu, Yaowei Wang, Zhaozhi Wang, Huizhu Jia 等AAAI 2021 · 被引用 88 次
相关 Paper
- FedLight: Federated Reinforcement Learning for Autonomous Multi-Intersection Traffic Signal ControlYutong Ye, Wupan Zhao, Tongquan Wei, Shiyan Hu 等DAC 2021 · 被引用 26 次
- CoSLight: Co-optimizing Collaborator Selection and Decision-making to Enhance Traffic Signal ControlJingqing Ruan, Ziyue Li, Hua Wei, Haoyuan Jiang 等KDD 2024 · 被引用 18 次
- Mitigating Action Hysteresis in Traffic Signal Control with Traffic Predictive Reinforcement LearningXiao Han, Xiangyu Zhao, Liang Zhang, Wanyu WangKDD 2023 · 被引用 16 次
- CrossLight: Offline-to-Online Reinforcement Learning for Cross-City Traffic Signal ControlQian Sun, Rui Zha, Le Zhang, Jingbo Zhou 等KDD 2024 · 被引用 9 次
- Optimizing Traffic Control with Model-Based Learning: A Pessimistic Approach to Data-Efficient Policy InferenceMayuresh Kunjir, Sanjay Chawla, Siddarth Chandrasekar, Devika Jay 等KDD 2023 · 被引用 3 次
