Hierarchically and Cooperatively Learning Traffic Signal Control
Bingyu Xu, Yaowei Wang, Zhaozhi Wang, Huizhu Jia, Zongqing Lu
Abstract
Deep reinforcement learning (RL) has been applied to traffic signal control recently and demonstrated superior performance to conventional control methods. However, there are still several challenges we have to address before fully applying deep RL to traffic signal control. Firstly, the objective of traffic signal control is to optimize average travel time, which is a delayed reward in a long time horizon in the context of RL. However, existing work simplifies the optimization by using queue length, waiting time, delay, etc., as immediate reward and presumes these short-term targets are always aligned with the objective. Nevertheless, these targets may deviate from the objective in different road networks with various traffic patterns. Secondly, it remains unsolved how to cooperatively control traffic signals to directly optimize average travel time. To address these challenges, we propose a hierarchical and cooperative reinforcement learning method-HiLight. HiLight enables each agent to learn a high-level policy that optimizes the objective locally by selecting among the sub-policies that respectively optimize short-term targets. Moreover, the high-level policy additionally considers the objective in the neighborhood with adaptive weighting to encourage agents to cooperate on the objective in the road network. Empirically, we demonstrate that HiLight outperforms state-of-the-art RL methods for traffic signal control in real road networks with real traffic.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers11
- FOP: Factorizing Optimal Joint Policy of Maximum-Entropy Multi-Agent Reinforcement LearningTianhao Zhang, Yueheng Li, Chen Wang, Guangming Xie et al.ICML 2021 · 88 citations
- OAM: An Option-Action Reinforcement Learning Framework for Universal Multi-Intersection ControlEnming Liang, Zicheng Su, Chilin Fang, Renxin ZhongAAAI 2022 · 30 citations
- EMVLight: A Decentralized Reinforcement Learning Framework for Efficient Passage of Emergency VehiclesHaoran Su, Yaofeng Desmond Zhong, Biswadip Dey, Amit ChakrabortyAAAI 2022 · 29 citations
- Difference Advantage Estimation for Multi-Agent Policy GradientsYueheng Li, Guangming Xie, Zongqing LuICML 2022 · 24 citations
- Two Heads are Better Than One: A Simple Exploration Framework for Efficient Multi-Agent Reinforcement LearningJiahui Li, Kun Kuang, Baoxiang Wang, Xingchen Li et al.NeurIPS 2023 · 7 citations
Builds on3
- Toward A Thousand Lights: Decentralized Deep Reinforcement Learning for Large-Scale Traffic Signal ControlChacha Chen, Hua Wei, Nan Xu, Guanjie Zheng et al.AAAI 2020 · 450 citations
- Graph Convolutional Reinforcement LearningJiechuan Jiang, Chen Dun, Tiejun Huang, Zongqing LuICLR 2020 · 415 citations
- MetaLight: Value-Based Meta-Reinforcement Learning for Traffic Signal ControlXinshi Zang, Huaxiu Yao, Guanjie Zheng, Nan Xu et al.AAAI 2020 · 185 citations
Related papers
- FedLight: Federated Reinforcement Learning for Autonomous Multi-Intersection Traffic Signal ControlYutong Ye, Wupan Zhao, Tongquan Wei, Shiyan Hu et al.DAC 2021 · 26 citations
- CoLLMLight: Cooperative Large Language Model Agents for Network-Wide Traffic Signal ControlZirui Yuan, Siqi Lai, Hao LiuICLR 2026 · 18 citations
- AttendLight: Universal Attention-Based Reinforcement Learning Model for Traffic Signal ControlAfshin Oroojlooy, MohammadReza Nazari, Davood Hajinezhad, Jorge SilvaNeurIPS 2020 · 143 citations
- CoSLight: Co-optimizing Collaborator Selection and Decision-making to Enhance Traffic Signal ControlJingqing Ruan, Ziyue Li, Hua Wei, Haoyuan Jiang et al.KDD 2024 · 18 citations
- HALO: Hierarchical Reinforcement Learning for Large-Scale Adaptive Traffic Signal ControlYaqiao Zhu, Hongkai Wen, Geyong Min, Man LuoWWW 2026
