Solving Hierarchical Information-Sharing Dec-POMDPs: An Extensive-Form Game Approach
Johan Peralez, Aurélien Delage, Olivier Buffet, Jilles Steeve Dibangoye
Abstract
A recent theory shows that a multi-player decentralized partially observable Markov decision process can be transformed into an equivalent single-player game, enabling the application of 's principle of optimality to solve the single-player game by breaking it down into single-stage subgames. However, this approach entangles the decision variables of all players at each single-stage subgame, resulting in backups with a double-exponential complexity. This paper demonstrates how to disentangle these decision variables while maintaining optimality under hierarchical information sharing, a prominent management style in our society. To achieve this, we apply the principle of optimality to solve any single-stage subgame by breaking it down further into smaller subgames, enabling us to make single-player decisions at a time. Our approach reveals that extensive-form games always exist with solutions to a single-stage subgame, significantly reducing time complexity. Our experimental results show that the algorithms leveraging these findings can scale up to much larger multi-player games without compromising optimality.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b5d419d8-e1ce-4aa4-9473-ad901fe367a8Cited by top-tier papers2
- Optimally Solving Simultaneous-Move Dec-POMDPs: The Sequential Central Planning ApproachJohan Peralez, Aurélien Delage, Jacopo Castellini, Rafael F. Cunha et al.AAAI 2025 · 3 citations
- ε-Optimally Solving Two-Player Zero-Sum POSGsErwan Escudie, Matthia Sabatelli, Olivier Buffet, Jilles DibangoyeNeurIPS 2025 · 1 citation
Builds on1
Related papers
- Efficient Multiagent Planning via Shared Action SuggestionsDylan M. Asmar, Mykel J. KochenderferAAAI 2026
- Decentralized Single-Timescale Actor-Critic on Zero-Sum Two-Player Stochastic GamesHongyi Guo, Zuyue Fu, Zhuoran Yang, Zhaoran WangICML 2021 · 11 citations
- Iterated Reasoning with Mutual Information in Cooperative and Byzantine Decentralized TeamingSachin G. Konan, Esmaeil Seraj, Matthew C. GombolayICLR 2022 · 27 citations
- Provably efficient multi-task reinforcement learning with model transferChicheng Zhang, Zhi WangNeurIPS 2021 · 20 citations
- Welfare Maximization in Competitive Equilibrium: Reinforcement Learning for Markov Exchange EconomyZhihan Liu, Miao Lu, Zhaoran Wang, Michael I. Jordan et al.ICML 2022 · 23 citations
