HyPOLE: Hyperproperty-Guided Multi-Agent Reinforcement Learning under Partial Observation
Arshia Rafieioskouei, Tzu-Han Hsu, Matthew Lucas, Borzoo Bonakdarpour
Abstract
Formal specification is a powerful tool to guide the learning process and provides significant advantages over reward shaping: (1) mathematical rigor; (2) expressiveness to specify objectives and constraints, and (3) the ability to define tactics to achieve objectives. However, these benefits remain largely unexplored in the context of Multi-Agent Reinforcement Learning (MARL). This paper introduces HyPOLE, a novel framework for MARL under partial observability, where learning is guided by the expressive power of the so-called hyperproperties and, in particular, the temporal logic HyperLTL. We integrate Centralized Training for Decentralized Execution (CTDE) techniques with HyPOLE to synthesize decentralized policies, and our evaluation on SMAC, MessySMAC, and WildFire benchmark demonstrates clear advantages over baselines.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 2bdf436a-b130-4e5d-a9b7-b68540c7954bBuilds on4
- Compositional Reinforcement Learning from Logical SpecificationsKishor Jothimurugan, Suguman Bansal, Osbert Bastani, Rajeev AlurNeurIPS 2021 · 112 citations
- Attention-Based Recurrence for Multi-Agent Reinforcement Learning under Stochastic Partial ObservabilityThomy Phan, Fabian Ritz, Philipp Altmann, Maximilian Zorn et al.ICML 2023 · 25 citations
- Specification-Guided Learning of Nash Equilibria with High Social WelfareKishor Jothimurugan, Suguman Bansal, Osbert Bastani, Rajeev AlurCAV 2022 · 9 citations
- HypRL: Reinforcement Learning of Control Policies for HyperpropertiesTzu-Han Hsu, Arshia Rafieioskouei, Borzoo BonakdarpourNeurIPS 2025 · 5 citations
Related papers
- Multi-Agent Guided Policy OptimizationYueheng Li, Guangming Xie, Zongqing LuICLR 2026 · 4 citations
- Decomposing Temporal Equilibrium Strategy for Coordinated Distributed Multi-Agent Reinforcement LearningChenyang Zhu, Wen Si, Jinyu Zhu, Zhihao JiangAAAI 2024 · 2 citations
- Enhancing Cooperative Multi-Agent Reinforcement Learning with State Modelling and Adversarial ExplorationAndreas Kontogiannis, Konstantinos Papathanasiou, Yi Shen, Giorgos Stamou et al.ICML 2025
- On the Expressivity of Objective-Specification Formalisms in Reinforcement LearningRohan Subramani, Marcus Williams, Max Heitmann, Halfdan Holm et al.ICLR 2024 · 3 citations
- Self-Organized Group for Cooperative Multi-agent Reinforcement LearningJianzhun Shao, Zhiqiang Lou, Hongchang Zhang, Yuhang Jiang et al.NeurIPS 2022 · 41 citations
