DiffLight: A Partial Rewards Conditioned Diffusion Model for Traffic Signal Control with Missing Data
Hanyang Chen, Yang Jiang, Shengnan Guo, Xiaowei Mao, Youfang Lin, Huaiyu Wan
Abstract
The application of reinforcement learning in traffic signal control (TSC) has been extensively researched and yielded notable achievements. However, most existing works for TSC assume that traffic data from all surrounding intersections is fully and continuously available through sensors. In real-world applications, this assumption often fails due to sensor malfunctions or data loss, making TSC with missing data a critical challenge. To meet the needs of practical applications, we introduce DiffLight, a novel conditional diffusion model for TSC under data-missing scenarios in the offline setting. Specifically, we integrate two essential sub-tasks, i.e., traffic data imputation and decision-making, by leveraging a Partial Rewards Conditioned Diffusion (PRCD) model to prevent missing rewards from interfering with the learning process. Meanwhile, to effectively capture the spatial-temporal dependencies among intersections, we design a Spatial-Temporal transFormer (STFormer) architecture. In addition, we propose a Diffusion Communication Mechanism (DCM) to promote better communication and control performance under data-missing scenarios. Extensive experiments on five datasets with various data-missing scenarios demonstrate that DiffLight is an effective controller to address TSC with missing data. The code of DiffLight is released at https://github.com/lokol5579/DiffLight-release.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 22105982-9b28-45ef-9f18-9d60acbd2896Cited by top-tier papers3
- VLMLight: Safety-Critical Traffic Signal Control via Vision-Language Meta-Control and Dual-Branch Reasoning ArchitectureMaonan Wang, Yirong Chen, Aoyu Pang, Yuxin Cai et al.NeurIPS 2025 · 6 citations
- Incomplete Data, Complete Dynamics: A Diffusion ApproachZihan Zhou, Chenguang Wang, Hongyi Ye, Yongtao Guan et al.ICLR 2026 · 3 citations
- RobustLight: Improving Robustness via Diffusion Reinforcement Learning for Traffic Signal ControlMingyuan Li, Jiahao Wang, Guangsheng Yu, Xu Wang et al.ICML 2025
Builds on20
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 11,743 citations
- Improved Denoising Diffusion Probabilistic ModelsAlexander Quinn Nichol, Prafulla DhariwalICML 2021 · 5,234 citations
- Conservative Q-Learning for Offline Reinforcement LearningAviral Kumar, Aurick Zhou, George Tucker, Sergey LevineNeurIPS 2020 · 2,881 citations
Related papers
- TransformerLight: A Novel Sequence Modeling Based Traffic Signaling Mechanism via Gated TransformerQiang Wu, Mingyuan Li, Jun Shen, Linyuan Lü et al.KDD 2023 · 17 citations
- PriSTI: A Conditional Diffusion Framework for Spatiotemporal ImputationMingzhe Liu, Han Huang, Hao Feng, Leilei Sun et al.ICDE 2023 · 110 citations
- Self-Supervised Learning of Time Series Representation via Diffusion Process and Imputation-Interpolation-Forecasting MaskZineb Senane, Lele Cao, Valentin Leonhard Buchner, Yusuke Tashiro et al.KDD 2024 · 17 citations
- RDPI: A Refine Diffusion Probability Generation Method for Spatiotemporal Data ImputationZijin Liu, Xiang Zhao, You SongAAAI 2025
- Reinformer: Max-Return Sequence Modeling for Offline RLZifeng Zhuang, Dengyun Peng, Jinxin Liu, Ziqi Zhang et al.ICML 2024 · 29 citations
