TransformerLight: A Novel Sequence Modeling Based Traffic Signaling Mechanism via Gated Transformer
Qiang Wu, Mingyuan Li, Jun Shen, Linyuan Lü, Bo Du, Ke Zhang
Abstract
Traffic signal control (TSC) is still one of the most significant and challenging research problems in the transportation field. Reinforcement learning (RL) has achieved great success in TSC but suffers from critically high learning costs in practical applications due to the excessive trial-and-error learning process. Offline RL is a promising method to reduce learning costs whereas the data distribution shift issue is still up in the air. To this end, in this paper, we formulate TSC as a sequence modeling problem with a sequence of Markov decision process described by states, actions, and rewards from the traffic environment. A novel framework, namely TransformerLight, is introduced, which does not aim to fit into value functions by averaging all possible returns, but produces the best possible actions using a gated Transformer. Additionally, the learning process of TransformerLight is much more stable by replacing the residual connections with gated transformer blocks due to a dynamic system perspective. Through numerical experiments on offline datasets, we demonstrate that the TransformerLight model: (1) can build a high-performance adaptive TSC model without dynamic programming; (2) achieves a new state-of-the-art compared to most published offline RL methods so far; and (3) shows a more stable learning process than offline RL and recent Transformer-based methods. The relevant dataset and code are available at Github.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get ece58d20-359b-4643-8154-fee5ed4cf924Cited by top-tier papers4
- CoLLMLight: Cooperative Large Language Model Agents for Network-Wide Traffic Signal ControlZirui Yuan, Siqi Lai, Hao LiuICLR 2026 · 18 citations
- CrossLight: Offline-to-Online Reinforcement Learning for Cross-City Traffic Signal ControlQian Sun, Rui Zha, Le Zhang, Jingbo Zhou et al.KDD 2024 · 9 citations
- CFLight: Enhancing Safety with Traffic Signal Control through Counterfactual LearningMingyuan Li, Chunyu Liu, Zhuojun Li, Xiao Liu et al.KDD 2026
- Scalable Traffic Signal Control with Shared Policy FrameworkHaolun MA, Yanchen ZHU, Zizhuo Xu, Weijie Shi et al.ICML 2026
Related papers
- Optimizing Traffic Control with Model-Based Learning: A Pessimistic Approach to Data-Efficient Policy InferenceMayuresh Kunjir, Sanjay Chawla, Siddarth Chandrasekar, Devika Jay et al.KDD 2023 · 3 citations
- Rethinking Decision Transformer via Hierarchical Reinforcement LearningYi Ma, Jianye Hao, Hebin Liang, Chenjun XiaoICML 2024 · 15 citations
- Bootstrapped Transformer for Offline Reinforcement LearningKerong Wang, Hanye Zhao, Xufang Luo, Kan Ren et al.NeurIPS 2022 · 54 citations
- Reinformer: Max-Return Sequence Modeling for Offline RLZifeng Zhuang, Dengyun Peng, Jinxin Liu, Ziqi Zhang et al.ICML 2024 · 29 citations
- DiffLight: A Partial Rewards Conditioned Diffusion Model for Traffic Signal Control with Missing DataHanyang Chen, Yang Jiang, Shengnan Guo, Xiaowei Mao et al.NeurIPS 2024 · 18 citations
