MARLIN: Multi-Agent Reinforcement Learning for Incremental DAG Discovery
Dong Li, Zhengzhang Chen, Xujiang Zhao, Linlin Yu, Zhong Chen, Yi He, Haifeng Chen, Chen Zhao
摘要
Uncovering causal structures from observational data is crucial for understanding complex systems and making informed decisions. While reinforcement learning (RL) has shown promise in identifying these structures in the form of a directed acyclic graph (DAG), existing methods often lack efficiency, making them unsuitable for online applications. In this paper, we propose MARLIN, an efficient multi-agent RL-based approach for incremental DAG learning. MAR-LIN uses a DAG generation policy that maps a continuous real-valued space to the DAG space as an intra-batch strategy, then incorporates two RL agents-state-specific and state-invariant-to uncover causal relationships and integrates these agents into an incremental learning framework. Furthermore, the framework leverages a factored action space to enhance parallelization efficiency. Extensive experiments on synthetic and real datasets demonstrate that MARLIN outperforms state-of-the-art methods in terms of both efficiency and effectiveness.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper10
- Causal Discovery with Reinforcement LearningShengyu Zhu, Ignavier Ng, Zhitang ChenICLR 2020 · 被引用 285 次
- Nezha: Interpretable Fine-Grained Root Causes Analysis for Microservices on Multi-modal Observability DataGuangba Yu, Pengfei Chen, Yufeng Li, Hongyang Chen 等FSE 2023 · 被引用 131 次
- Leveraging Factored Action Spaces for Efficient Offline Reinforcement Learning in HealthcareShengpu Tang, Maggie Makar, Michael W. Sjoding, Finale Doshi-Velez 等NeurIPS 2022 · 被引用 63 次
- MULAN: Multi-modal Causal Structure Learning and Root Cause Analysis for Microservice SystemsLecheng Zheng, Zhengzhang Chen, Jingrui He, Haifeng ChenWWW 2024 · 被引用 53 次
- Differentiable DAG SamplingBertrand Charpentier, Simon Kibler, Stephan GünnemannICLR 2022 · 被引用 51 次
相关 Paper
- Reinforcement Causal Structure Learning on Order GraphDezhi Yang, Guoxian Yu, Jun Wang, Zhengtian Wu 等AAAI 2023 · 被引用 20 次
- IDYNO: Learning Nonparametric DAGs from Interventional Dynamic DataTian Gao, Debarun Bhattacharjya, Elliot Nelson, Miao Liu 等ICML 2022 · 被引用 26 次
- Hierarchical Reinforcement Learning with Targeted Causal InterventionsMohammadsadegh Khorasani, Saber Salehkaleybar, Negar Kiyavash, Matthias GrossglauserICML 2025
- Meta-D2AG: Causal Graph Learning with Interventional Dynamic DataTian Gao, Songtao Lu, Junkyu Lee, Elliot Nelson 等NeurIPS 2025
- Gradient-Based Neural DAG LearningSébastien Lachapelle, Philippe Brouillard, Tristan Deleu, Simon Lacoste-JulienICLR 2020 · 被引用 337 次
