Secure Multi-agent Reinforcement Learning for Service Systems with Affinity and Byzantine Nodes: Stability Analysis and Protection Design
Yifan Jiang, Jiasheng Pan, Mengtian Li, Li Jin
Abstract
We study decentralized multi-agent reinforcement learning (MARL) for networked service systems with affinity in the presence of Byzantine nodes. The way that a server processes a job depends on an affinity state that captures the correlation between the job and the server. Each node learns a local control policy via an actor-critic algorithm with linear function approximation over inherently unbounded space of traffic states, while exchanging parameter information with neighbors through a communication graph. A set of Byzantine agents can exploit the unbounded state space to compromise the consensus mechanism, destabilizing both learning and queuing processes. To address this vulnerability, we propose a resilient consensus-based MARL algorithm, which mitigates adversarial parameter manipulation and guarantees traffic stability under mild assumptions. We prove that the cooperative agents’ policies converge almost surely to a bounded neighborhood of a stationary solution of the global objective. We demonstrate the effectiveness and generality of the proposed framework in several representative service systems, including semantic routing for large language model serving, distributed polling in cloud computing, and smart manufacturing logistics.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b304b5e3-afa0-4d3e-8417-0ef6bc69d535Builds on3
- Toward A Thousand Lights: Decentralized Deep Reinforcement Learning for Large-Scale Traffic Signal ControlChacha Chen, Hua Wei, Nan Xu, Guanjie Zheng et al.AAAI 2020 · 450 citations
- Multi-Agent Reinforcement Learning in Stochastic Networked SystemsYiheng Lin, Guannan Qu, Longbo Huang, Adam WiermanNeurIPS 2021 · 55 citations
- Byzantine Robust Cooperative Multi-Agent Reinforcement Learning as a Bayesian GameSimin Li, Jun Guo, Jingqiao Xiu, Ruixiao Xu et al.ICLR 2024 · 30 citations
Related papers
- Reinforcement Learning Based Multi-Agent Resilient Control: From Deep Neural Networks to an Adaptive LawJian Hou, Fangyuan Wang, Lili Wang, Zhiyong ChenAAAI 2021 · 7 citations
- A Universal Transcoding and Transmission Method for Livecast with Networked Multi-Agent Reinforcement LearningXingyan Chen, Changqiao Xu, Mu Wang, Zhonghui Wu et al.INFOCOM 2021 · 16 citations
- BR-DeFedRL: Byzantine-Robust Decentralized Federated Reinforcement Learning with Fast Convergence and Communication EfficiencyJing Qiao, Zuyuan Zhang, Sheng Yue, Yuan Yuan et al.INFOCOM 2024 · 13 citations
- Byzantine Resilient Distributed Multi-Task LearningJiani Li, Waseem Abbas, Xenofon D. KoutsoukosNeurIPS 2020 · 12 citations
- Flag Aggregator: Scalable Distributed Training under Failures and Augmented Losses using Convex OptimizationHamidreza Almasi, Harsh Mishra, Balajee Vamanan, Sathya N. RaviICLR 2024
