HMARL-CBF - Hierarchical Multi-Agent Reinforcement Learning with Control Barrier Functions for Safety-Critical Autonomous Systems
H. M. Sabbir Ahmad, Ehsan Sabouni, Alexander Wasilkoff, Param Budhraja, Zijian Guo, Songyuan Zhang, Chuchu Fan, Christos G. Cassandras, Wenchao Li
摘要
We address the problem of safe policy learning in multi-agent safety-critical autonomous systems. In such systems, it is necessary for each agent to meet the safety requirements at all times while also cooperating with other agents to accomplish the task. Toward this end, we propose a safe Hierarchical Multi-Agent Reinforcement Learning (HMARL) approach based on Control Barrier Functions (CBFs). Our proposed hierarchical approach decomposes the overall reinforcement learning problem into two levels -learning joint cooperative behavior at the higher level and learning safe individual behavior at the lower or agent level, conditioned on the high-level policy. Specifically, we propose a skill-based HMARL-CBF algorithm in which the higher-level problem involves learning a joint policy over the skills for all the agents, and the lower-level problem involves learning policies to execute the skills safely with CBFs. We validate our approach in challenging environment scenarios, whereby a large number of agents have to safely navigate through conflicting road networks. Compared with existing state-of-the-art methods, our approach significantly improves the safety, achieving a near-perfect (≥ 95%) success/safety rate while improving performance across all the environments 1 .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper20
- Learning Safe Multi-agent Control with Decentralized Neural Barrier CertificatesZengyi Qin, Kaiqing Zhang, Yuxiao Chen, Jingkai Chen 等ICLR 2021 · 被引用 164 次
- Multi-Robot Collision Avoidance under Uncertainty with Probabilistic Safety Barrier CertificatesWenhao Luo, Wen Sun, Ashish KapoorNeurIPS 2020 · 被引用 102 次
- Reachability Constrained Reinforcement LearningDongjie Yu, Haitong Ma, Sheng-bo Li, Jianyu ChenICML 2022 · 被引用 90 次
- Learning to Simulate Self-driven Particles System with Coordinated Policy OptimizationZhenghao Peng, Quanyi Li, Ka-Ming Hui, Chunxiao Liu 等NeurIPS 2021 · 被引用 88 次
- Enforcing Hard Constraints with Soft Barriers: Safe Reinforcement Learning in Unknown Stochastic EnvironmentsYixuan Wang, Simon Sinong Zhan, Ruochen Jiao, Zhilu Wang 等ICML 2023 · 被引用 81 次
相关 Paper
- A Physics-Informed Machine Learning Framework for Safe and Optimal Control of Autonomous SystemsManan Tayal, Aditya Singh, Shishir Kolathaya, Somil BansalICML 2025
- Hierarchical Multi-Agent Skill DiscoveryMingyu Yang, Yaodong Yang, Zhenbo Lu, Wengang Zhou 等NeurIPS 2023 · 被引用 34 次
- Learning Verified Safe Neural Network Controllers for Multi-Agent Path FindingMingyue Zhang, Nianyu Li, Yi Chen, Jialong Li 等AAAI 2025 · 被引用 2 次
- Multi-Agent First Order Constrained Optimization in Policy SpaceYoupeng Zhao, Yaodong Yang, Zhenbo Lu, Wengang Zhou 等NeurIPS 2023 · 被引用 12 次
- Scalable Constrained Policy Optimization for Safe Multi-agent Reinforcement LearningLijun Zhang, Lin Li, Wei Wei, Huizhong Song 等NeurIPS 2024 · 被引用 22 次
