Integrating Suboptimal Human Knowledge with Hierarchical Reinforcement Learning for Large-Scale Multiagent Systems
Dingbang Liu, Shohei Kato, Wen Gu, Fenghui Ren, Jun Yan, Guoxin Su
Abstract
Due to the exponential growth of agent interactions and the curse of dimensionality, learning efficient coordination from scratch is inherently challenging in large-scale multi-agent systems. While agents’ learning is data-driven, sampling from millions of steps, human learning processes are quite different. Inspired by the concept of Human-on-the-Loop and the daily human hierarchical control, we propose a novel knowledge-guided multi-agent reinforcement learning framework (hhk-MARL), which combines human abstract knowledge with hierarchical reinforcement learning to address the learning difficulties among a large number of agents. In this work, fuzzy logic is applied to represent human suboptimal knowledge, and agents are allowed to freely decide how to leverage the proposed prior knowledge. Additionally, a graph-based group controller is built to enhance agent coordination. The proposed framework is end-to-end and compatible with various existing algorithms. We conduct experiments in challenging domains of the StarCraft Multi-agent Challenge combined with three famous algorithms: IQL, QMIX, and Qatten. The results show that our approach can greatly accelerate the training process and improve the final performance, even based on low-performance human prior knowledge.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext f36a7478-bd78-4f59-b710-5a314a30fb14Cited by top-tier papers3
- Unlocking the Power of Multi-Agent LLM for Reasoning: From Lazy Agents to DeliberationZhiwei Zhang, Xiaomin Li, Yudi Lin, Hui Liu et al.ICLR 2026 · 13 citations
- A Principle of Targeted Intervention for Multi-Agent Reinforcement LearningAnjie Liu, Jianhong Wang, Samuel Kaski, Jun Wang et al.NeurIPS 2025 · 3 citations
- Strategic Planning: A Top-Down Approach to Option GenerationMax Ruiz Luyten, Antonin Berthon, Mihaela van der SchaarICML 2025
Builds on5
- Multi-Agent Game Abstraction via Graph Attention Neural NetworkYong Liu, Weixun Wang, Yujing Hu, Jianye Hao et al.AAAI 2020 · 316 citations
- Scaling Multi-Agent Reinforcement Learning with Selective Parameter SharingFilippos Christianos, Georgios Papoudakis, Arrasy Rahman, Stefano V. AlbrechtICML 2021 · 165 citations
- From Few to More: Large-Scale Dynamic Multiagent Curriculum LearningWeixun Wang, Tianpei Yang, Yong Liu, Jianye Hao et al.AAAI 2020 · 138 citations
- Hierarchical Mean-Field Deep Reinforcement Learning for Large-Scale Multiagent SystemsChao YuAAAI 2023 · 8 citations
- Evolutionary Population Curriculum for Scaling Multi-Agent Reinforcement LearningQian Long, Zihan Zhou, Abhinav Gupta, Fei Fang et al.ICLR 2020
Related papers
- Reinforcement Learning with Fuzzy Human Attention-Guided Graph for Heterogeneous Multiagent SystemsDingbang Liu, Fenghui Ren, Jun Yan, Guoxin Su et al.AAAI 2026
- Automatic Grouping for Efficient Cooperative Multi-Agent Reinforcement LearningYifan Zang, Jinmin He, Kai Li, Haobo Fu et al.NeurIPS 2023 · 37 citations
- Logic-Q: Improving Deep Reinforcement Learning-based Quantitative Trading via Program Sketch-based TuningZhiming Li, Junzhe Jiang, Yushi Cao, Aixin Cui et al.AAAI 2025 · 5 citations
- ROMA: Multi-Agent Reinforcement Learning with Emergent RolesTonghan Wang, Heng Dong, Victor R. Lesser, Chongjie ZhangICML 2020 · 286 citations
- HCPO: Hierarchical Conductor-Based Policy Optimization in Multi-Agent Reinforcement LearningZejiao Liu, Junqi Tu, Yitian Hong, Luolin Xiong et al.AAAI 2026
