DAGMA: Learning DAGs via M-matrices and a Log-Determinant Acyclicity Characterization
Kevin Bello, Bryon Aragam, Pradeep Ravikumar
摘要
The combinatorial problem of learning directed acyclic graphs (DAGs) from data was recently framed as a purely continuous optimization problem by leveraging a differentiable acyclicity characterization of DAGs based on the trace of a matrix exponential function. Existing acyclicity characterizations are based on the idea that powers of an adjacency matrix contain information about walks and cycles. In this work, we propose a new acyclicity characterization based on the log-determinant (log-det) function, which leverages the nilpotency property of DAGs. To deal with the inherent asymmetries of a DAG, we relate the domain of our log-det characterization to the set of , which is a key difference to the classical log-det function defined over the cone of positive definite matrices. Similar to acyclicity functions previously proposed, our characterization is also exact and differentiable. However, when compared to existing characterizations, our log-det function: (1) Is better at detecting large cycles; (2) Has better-behaved gradients; and (3) Its runtime is in practice about an order of magnitude faster. From the optimization side, we drop the typically used augmented Lagrangian scheme and propose DAGMA (), a method that resembles the central path for barrier methods. Each point in the central path of DAGMA is a solution to an unconstrained problem regularized by our log-det function, then we show that at the limit of the central path the solution is guaranteed to be a DAG. Finally, we provide extensive experiments for and SEMs and show that our approach can reach large speed-ups and smaller structural Hamming distances against state-of-the-art methods. Code implementing the proposed method is open-source and publicly available at https://github.com/kevinsbello/dagma.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper43
- Learning Linear Causal Representations from Interventions under General Nonlinear MixingSimon Buchholz, Goutham Rajendran, Elan Rosenfeld, Bryon Aragam 等NeurIPS 2023 · 被引用 113 次
- Fast Scalable and Accurate Discovery of DAGs Using the Best Order Score Search and Grow Shrink TreesBryan Andrews, Joseph D. Ramsey, Ruben Sanchez-Romero, Jazmin Camchong 等NeurIPS 2023 · 被引用 61 次
- Stable Differentiable Causal DiscoveryAchille Nazaret, Justin Hong, Elham Azizi, David M. BleiICML 2024 · 被引用 29 次
- Optimizing NOTEARS Objectives via Topological SwapsChang Deng, Kevin Bello, Bryon Aragam, Pradeep Kumar RavikumarICML 2023 · 被引用 23 次
- Learning DAGs from Data with Few Root CausesPanagiotis Misiakos, Chris Wendler, Markus PüschelNeurIPS 2023 · 被引用 17 次
它引用的顶会 Paper4
- On the Role of Sparsity and DAG Constraints for Learning Linear DAGsIgnavier Ng, AmirEmad Ghassami, Kun ZhangNeurIPS 2020 · 被引用 306 次
- DAGs with No Fears: A Closer Look at Continuous Optimization for Learning Bayesian NetworksDennis Wei, Tian Gao, Yue YuNeurIPS 2020 · 被引用 102 次
- CASTLE: Regularization via Auxiliary Causal Graph DiscoveryTrent Kyono, Yao Zhang, Mihaela van der SchaarNeurIPS 2020 · 被引用 82 次
- DAGs with No Curl: An Efficient DAG Structure Learning ApproachYue Yu, Tian Gao, Naiyu Yin, Qiang JiICML 2021 · 被引用 77 次
相关 Paper
- Constraint-Free Structure Learning with Smooth Acyclic OrientationsRiccardo Massidda, Francesco Landolfi, Martina Cinquini, Davide BacciuICLR 2024 · 被引用 10 次
- Analytic DAG Constraints for Differentiable DAG LearningZhen Zhang, Ignavier Ng, Dong Gong, Yuhang Liu 等ICLR 2025
- Truncated Matrix Power Iteration for Differentiable DAG LearningZhen Zhang, Ignavier Ng, Dong Gong, Yuhang Liu 等NeurIPS 2022 · 被引用 36 次
- CoLiDE: Concomitant Linear DAG EstimationSeyed Saman Saboksayr, Gonzalo Mateos, Mariano TepperICLR 2024 · 被引用 9 次
- Markov Equivalence and Consistency in Differentiable Structure LearningChang Deng, Kevin Bello, Pradeep Ravikumar, Bryon AragamNeurIPS 2024 · 被引用 8 次
