Vulnerable Agent Identification in Large-Scale Multi-Agent Reinforcement Learning
Simin Li, Zihao Mao, Zheng Yuwei, Linhao Wang, Ruixiao Xu, Chengdong Ma, Zhiqian Liu, Xin Yu, Yuqing Ma, Xin Wang, Jie Luo, Bo An
Abstract
Partial agent failure becomes inevitable when systems scale up, making it crucial to identify the subset of agents whose failure causes worst-case system performance degradations. We study this Vulnerable Agent Identification (VAI) problem in large-scale multi-agent reinforcement learning (MARL). We frame VAI as a Hierarchical Adversarial Decentralized Mean Field Control (HAD-MFC), where where the upper level selects vulnerable agents as an NP-hard task and the lower level learns their worst-case adversarial policies via mean-field MARL. The two problems are coupled together, making HAD-MFC difficult to solve. To handle this, we first decouple the hierarchical process by Fenchel-Rockafellar transform, resulting a regularized mean-field Bellman operator for upper level that enables independent learning at each level, thus reducing computational complexity. We next reformulate the upper-level NP-hard problem as an MDP with dense rewards, allowing sequential identification of vulnerable agents via greedy and RL algorithms. This decomposition provably preserves the optimal solution. Experiments show our method effectively identifies more vulnerable agents in large-scale MARL and the rule-based system, fooling system into worse failures, and reveals the vulnerability of each agent in large systems. Code available at https://anonymous.4open.science/r/VAI-5F61/.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on12
- Adversarial Policies: Attacking Deep Reinforcement LearningAdam Gleave, Michael Dennis, Cody Wild, Neel Kant et al.ICLR 2020 · 415 citations
- Multi-Agent Reinforcement Learning for Active Voltage Control on Power Distribution NetworksJianhong Wang, Wangkun Xu, Yunjie Gu, Wenbin Song et al.NeurIPS 2021 · 216 citations
- Deep Graph Representation Learning and Optimization for Influence MaximizationChen Ling, Junji Jiang, Junxiang Wang, My T. Thai et al.ICML 2023 · 159 citations
- Robust Multi-Agent Reinforcement Learning with Model UncertaintyKaiqing Zhang, Tao Sun, Yunzhe Tao, Sahika Genc et al.NeurIPS 2020 · 118 citations
- Scalable Deep Reinforcement Learning Algorithms for Mean Field GamesMathieu Laurière, Sarah Perrin, Sertan Girgin, Paul Muller et al.ICML 2022 · 64 citations
Related papers
- Major-Minor Mean Field Multi-Agent Reinforcement LearningKai Cui, Christian Fabian, Anam Tahir, Heinz KoepplICML 2024 · 6 citations
- Budget-Efficient Attacks and Robustness Training for Cooperative MARLJunyong Jiang, Xin Yuan, Longhe Lin, Songze Li et al.ICML 2026
- Hierarchical Mean-Field Deep Reinforcement Learning for Large-Scale Multiagent SystemsChao YuAAAI 2023 · 8 citations
- On Imitation in Mean-field GamesGiorgia Ramponi, Pavel Kolev, Olivier Pietquin, Niao He et al.NeurIPS 2023 · 12 citations
- Multi-Agent Imitation by Learning and Sampling from Factorized Soft Q-FunctionYi-Chen Li, Zhongxiang Ling, Tao Jiang, Fuxiang Zhang et al.NeurIPS 2025 · 3 citations
