CoLLMLight: Cooperative Large Language Model Agents for Network-Wide Traffic Signal Control
Zirui Yuan, Siqi Lai, Hao Liu
摘要
Large Language Models (LLMs) have recently emerged as promising agents for Traffic Signal Control (TSC) due to their strengths in reasoning and generalization. However, current LLM-based approaches treat intersections as independent agents without inter-intersection cooperation, limiting their effectiveness in network-wide optimization. To address this gap, we propose CoLLMLight, the first cooperative LLM agent framework for network-wide traffic signal control. CoLLMLight enables agents to perform in-depth spatiotemporal reasoning for cooperation, while ensuring real-time responsiveness through an asynchronous cooperative decision architecture. The reasoning process runs asynchronously, deriving cooperative control guidance from dynamic interactions among intersections. This guidance is cached and incorporated as contextual input for real-time signal decisions. To enhance cooperation quality while ensuring reasoning efficiency, we propose cost-aware cooperation optimization. It first applies adaptive reasoning chain optimization to enable the LLM to adjust its reasoning depth according to traffic complexity. The model is then refined with reinforcement learning using reward signals that promote network-wide performance while penalizing excessive reasoning. Extensive experiments on four real-world traffic networks demonstrate that CoLLMLight consistently outperforms existing methods, achieving more effective and generalizable cooperation while maintaining real-time responsiveness and efficient token usage.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- ARROW: An Adaptive Rollout and Routing Method for Global Weather ForecastingJindong Tian, Yifei Ding, Ronghui Xu, Hao Miao 等ICLR 2026 · 被引用 14 次
- Traffic-R1: Reinforced LLMs Bring Human-Like Reasoning to Traffic Signal Control SystemsXingchen Zou, Yuhao Yang, Zheng Chen, Xixuan Hao 等ACL 2026 · 被引用 9 次
- USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning Capabilities of LLMs as Urban AgentsSiqi Lai, Yansong Ning, Zirui Yuan, Zhixi Chen 等ICLR 2026 · 被引用 7 次
- An LLM-Powered Cooperative Framework for Large-Scale Multi-Vehicle NavigationYuping Zhou, Siqi Lai, Jindong Han, Hao LiuWWW 2026 · 被引用 2 次
- Towards Multimodal Data-Driven Scientific Discovery Powered by LLM AgentsFan Liu, Xiaozhao Zeng, Hao LiuICLR 2026
它引用的顶会 Paper13
- Chain-of-Thought Prompting Elicits Reasoning in Large Language ModelsJason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma 等NeurIPS 2022 · 被引用 22,562 次
- Generative Agents: Interactive Simulacra of Human BehaviorJoon Sung Park, Joseph C. O'Brien, Carrie Jun Cai, Meredith Ringel Morris 等UIST 2023 · 被引用 1,882 次
- Toward A Thousand Lights: Decentralized Deep Reinforcement Learning for Large-Scale Traffic Signal ControlChacha Chen, Hua Wei, Nan Xu, Guanjie Zheng 等AAAI 2020 · 被引用 450 次
- Building Cooperative Embodied Agents Modularly with Large Language ModelsHongxin Zhang, Weihua Du, Jiaming Shan, Qinhong Zhou 等ICLR 2024 · 被引用 303 次
- Deep Coordination GraphsWendelin Boehmer, Vitaly Kurin, Shimon WhitesonICML 2020 · 被引用 209 次
相关 Paper
- Orchestrating Reasoning and Reaction: An Asynchronous Hierarchical Framework for LLM-driven Traffic Signal ControlFansheng Sun, Jiyu Wang, Zhidan LiuKDD 2026
- VLMLight: Safety-Critical Traffic Signal Control via Vision-Language Meta-Control and Dual-Branch Reasoning ArchitectureMaonan Wang, Yirong Chen, Aoyu Pang, Yuxin Cai 等NeurIPS 2025 · 被引用 6 次
- Chain-of-Experts: When LLMs Meet Complex Operations Research ProblemsZiyang Xiao, Dongxiang Zhang, Yangjun Wu, Lilin Xu 等ICLR 2024 · 被引用 136 次
- Hierarchically and Cooperatively Learning Traffic Signal ControlBingyu Xu, Yaowei Wang, Zhaozhi Wang, Huizhu Jia 等AAAI 2021 · 被引用 88 次
- FedLight: Federated Reinforcement Learning for Autonomous Multi-Intersection Traffic Signal ControlYutong Ye, Wupan Zhao, Tongquan Wei, Shiyan Hu 等DAC 2021 · 被引用 26 次
