MetaLight: Value-Based Meta-Reinforcement Learning for Traffic Signal Control
Xinshi Zang, Huaxiu Yao, Guanjie Zheng, Nan Xu, Kai Xu, Zhenhui Li
Abstract
Using reinforcement learning for traffic signal control has attracted increasing interests recently. Various value-based reinforcement learning methods have been proposed to deal with this classical transportation problem and achieved better performances compared with traditional transportation methods. However, current reinforcement learning models rely on tremendous training data and computational resources, which may have bad consequences (e.g., traffic jams or accidents) in the real world. In traffic signal control, some algorithms have been proposed to empower quick learning from scratch, but little attention is paid to learning by transferring and reusing learned experience. In this paper, we propose a novel framework, named as MetaLight, to speed up the learning process in new scenarios by leveraging the knowledge learned from existing scenarios. MetaLight is a value-based meta-reinforcement learning workflow based on the representative gradient-based meta-learning algorithm (MAML), which includes periodically alternate individual-level adaptation and global-level adaptation. Moreover, MetaLight improves the-state-of-the-art reinforcement learning model FRAP in traffic signal control by optimizing its model structure and updating paradigm. The experiments on four real-world datasets show that our proposed MetaLight not only adapts more quickly and stably in new traffic scenarios, but also achieves better performance.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers10
- AttendLight: Universal Attention-Based Reinforcement Learning Model for Traffic Signal ControlAfshin Oroojlooy, MohammadReza Nazari, Davood Hajinezhad, Jorge SilvaNeurIPS 2020 · 143 citations
- Hierarchically and Cooperatively Learning Traffic Signal ControlBingyu Xu, Yaowei Wang, Zhaozhi Wang, Huizhu Jia et al.AAAI 2021 · 88 citations
- Prompt to Transfer: Sim-to-Real Transfer for Traffic Signal Control with Prompt LearningLongchao Da, Minquan Gao, Hao Mei, Hua WeiAAAI 2024 · 60 citations
- EMVLight: A Decentralized Reinforcement Learning Framework for Efficient Passage of Emergency VehiclesHaoran Su, Yaofeng Desmond Zhong, Biswadip Dey, Amit ChakrabortyAAAI 2022 · 29 citations
- DiffLight: A Partial Rewards Conditioned Diffusion Model for Traffic Signal Control with Missing DataHanyang Chen, Yang Jiang, Shengnan Guo, Xiaowei Mao et al.NeurIPS 2024 · 18 citations
Related papers
- CrossLight: Offline-to-Online Reinforcement Learning for Cross-City Traffic Signal ControlQian Sun, Rui Zha, Le Zhang, Jingbo Zhou et al.KDD 2024 · 9 citations
- FedLight: Federated Reinforcement Learning for Autonomous Multi-Intersection Traffic Signal ControlYutong Ye, Wupan Zhao, Tongquan Wei, Shiyan Hu et al.DAC 2021 · 26 citations
- Expression might be enough: representing pressure and demand for reinforcement learning based traffic signal controlLiang Zhang, Qiang Wu, Jun Shen, Linyuan Lü et al.ICML 2022 · 57 citations
- TransformerLight: A Novel Sequence Modeling Based Traffic Signaling Mechanism via Gated TransformerQiang Wu, Mingyuan Li, Jun Shen, Linyuan Lü et al.KDD 2023 · 17 citations
- Mitigating Action Hysteresis in Traffic Signal Control with Traffic Predictive Reinforcement LearningXiao Han, Xiangyu Zhao, Liang Zhang, Wanyu WangKDD 2023 · 16 citations
