On the Role of Sparsity and DAG Constraints for Learning Linear DAGs
Ignavier Ng, AmirEmad Ghassami, Kun Zhang
Abstract
Learning graphical structures based on Directed Acyclic Graphs (DAGs) is a challenging problem, partly owing to the large search space of possible graphs. A recent line of work formulates the structure learning problem as a continuous constrained optimization task using the least squares objective and an algebraic characterization of DAGs. However, the formulation requires a hard DAG constraint and may lead to optimization difficulties. In this paper, we study the asymptotic role of the sparsity and DAG constraints for learning DAG models in the linear Gaussian and non-Gaussian cases, and investigate their usefulness in the finite sample regime. Based on the theoretical results, we formulate a likelihood-based score function, and show that one only has to apply soft sparsity and DAG constraints to learn a DAG equivalent to the ground truth DAG. This leads to an unconstrained optimization problem that is much easier to solve. Using gradient-based optimization and GPU acceleration, our procedure can easily handle thousands of nodes while retaining a high accuracy. Extensive experiments validate the effectiveness of our proposed method and show that the DAG-penalized likelihood objective is indeed favorable over the least squares one with the hard DAG constraint.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers71
- DAGMA: Learning DAGs via M-matrices and a Log-Determinant Acyclicity CharacterizationKevin Bello, Bryon Aragam, Pradeep RavikumarNeurIPS 2022 · 222 citations
- Beware of the Simulated DAG! Causal Discovery Benchmarks May Be Easy to GameAlexander G. Reisach, Christof Seiler, Sebastian WeichwaldNeurIPS 2021 · 213 citations
- DiBS: Differentiable Bayesian Structure LearningLars Lorch, Jonas Rothfuss, Bernhard Schölkopf, Andreas KrauseNeurIPS 2021 · 144 citations
- BCD Nets: Scalable Variational Approaches for Bayesian Causal DiscoveryChris Cundy, Aditya Grover, Stefano ErmonNeurIPS 2021 · 105 citations
- Efficient Neural Causal Discovery without Acyclicity ConstraintsPhillip Lippe, Taco Cohen, Efstratios GavvesICLR 2022 · 95 citations
Builds on3
- Gradient-Based Neural DAG LearningSébastien Lachapelle, Philippe Brouillard, Tristan Deleu, Simon Lacoste-JulienICLR 2020 · 337 citations
- Causal Discovery with Reinforcement LearningShengyu Zhu, Ignavier Ng, Zhitang ChenICLR 2020 · 285 citations
- Characterizing Distribution Equivalence and Structure Learning for Cyclic and Acyclic Directed GraphsAmirEmad Ghassami, Alan Yang, Negar Kiyavash, Kun ZhangICML 2020 · 32 citations
Related papers
- Markov Equivalence and Consistency in Differentiable Structure LearningChang Deng, Kevin Bello, Pradeep Ravikumar, Bryon AragamNeurIPS 2024 · 8 citations
- DAGs with No Curl: An Efficient DAG Structure Learning ApproachYue Yu, Tian Gao, Naiyu Yin, Qiang JiICML 2021 · 77 citations
- CoLiDE: Concomitant Linear DAG EstimationSeyed Saman Saboksayr, Gonzalo Mateos, Mariano TepperICLR 2024 · 9 citations
- ProDAG: Projected Variational Inference for Directed Acyclic GraphsRyan Thompson, Edwin V. Bonilla, Robert KohnNeurIPS 2025 · 6 citations
- Integer Programming for Causal Structure Learning in the Presence of Latent VariablesRui Chen, Sanjeeb Dash, Tian GaoICML 2021 · 19 citations
