Local Policies for Graph-Structured Markov Decision Processes
Fathima Faizal, Asuman Ozdaglar, Martin Wainwright
摘要
We study a cooperative form of multi-agent reinforcement learning with state space dynamics and agent interaction controlled by an underlying graph. Each agent has a local state and action, the evolution of the local state depends only on the states and actions in the -hop neighborhood defined by the graph. Structured dynamics of this type arise in various applications, including network resource allocation, co-operative games, epidemic control, and wireless scheduling. The global state-action space scales exponentially in the number of agents, so that computing global optimal policies is intractable in the worst-case. We study conditions under which it is possible to approximate the optimal policies by a local policy for each agent that depends only on states associated with nodes within its -hop neighborhood. By controlling the propagation of influences via a Dobrushin-type stability matrix, we establish that globally optimal policies can approximated by local policies with sub-optimality gap decaying exponentially in .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper2
相关 Paper
- Multi-Agent Reinforcement Learning in Stochastic Networked SystemsYiheng Lin, Guannan Qu, Longbo Huang, Adam WiermanNeurIPS 2021 · 被引用 55 次
- Reinforcement Learning under a Multi-agent Predictive State Representation Model: Method and TheoryZhi Zhang, Zhuoran Yang, Han Liu, Pratap Tokekar 等ICLR 2022 · 被引用 10 次
- Emergent Fast-Slow Dynamics in Multi-Agent Q-Learning for Networked Stochastic GamesYuxin Geng, Wolfram Barfuss, Xingru ChenAAAI 2026
- Learning Graphon Mean Field Games and Approximate Nash EquilibriaKai Cui, Heinz KoepplICLR 2022 · 被引用 50 次
- Learning Mean Field Control on Sparse GraphsChristian Fabian, Kai Cui, Heinz KoepplICML 2025
