Revisiting Differentiable Structure Learning: Inconsistency of L1 Penalty and Beyond
Kaifeng Jin, Ignavier Ng, Kun Zhang, Biwei Huang
摘要
Recent advances in differentiable structure learning have framed the combinatorial problem of learning directed acyclic graphs as a continuous optimization problem. Various aspects, including data standardization, have been studied to identify factors that influence the empirical performance of these methods. In this work, we investigate critical limitations in differentiable structure learning methods, focusing on settings where the true structure can be identified up to Markov equivalence classes, particularly in the linear Gaussian case. While Ng et al. ( 2024 ) highlighted potential non-convexity issues in this setting, we demonstrate and explain why the use of ℓ 1 -penalized likelihood in such cases is fundamentally inconsistent, even if the global optimum of the optimization problem can be found. To resolve this limitation, we develop a hybrid differentiable structure learning method based on ℓ 0 -penalized likelihood with hard acyclicity constraint, where the ℓ 0 penalty can be approximated by different techniques including Gumbel-Softmax. Specifically, we first estimate the underlying moral graph, and use it to restrict the search space of the optimization problem, which helps alleviate the non-convexity issue. Experimental results show that the proposed method enhances empirical performance both before and after data standardization, providing a more reliable path for future advancements in differentiable structure learning, especially for learning Markov equivalence classes.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper11
- Gradient-Based Neural DAG LearningSébastien Lachapelle, Philippe Brouillard, Tristan Deleu, Simon Lacoste-JulienICLR 2020 · 被引用 337 次
- On the Role of Sparsity and DAG Constraints for Learning Linear DAGsIgnavier Ng, AmirEmad Ghassami, Kun ZhangNeurIPS 2020 · 被引用 306 次
- Differentiable Causal Discovery from Interventional DataPhilippe Brouillard, Sébastien Lachapelle, Alexandre Lacoste, Simon Lacoste-Julien 等NeurIPS 2020 · 被引用 295 次
- DAGMA: Learning DAGs via M-matrices and a Log-Determinant Acyclicity CharacterizationKevin Bello, Bryon Aragam, Pradeep RavikumarNeurIPS 2022 · 被引用 222 次
- Beware of the Simulated DAG! Causal Discovery Benchmarks May Be Easy to GameAlexander G. Reisach, Christof Seiler, Sebastian WeichwaldNeurIPS 2021 · 被引用 213 次
相关 Paper
- Markov Equivalence and Consistency in Differentiable Structure LearningChang Deng, Kevin Bello, Pradeep Ravikumar, Bryon AragamNeurIPS 2024 · 被引用 8 次
- Constraint-Free Structure Learning with Smooth Acyclic OrientationsRiccardo Massidda, Francesco Landolfi, Martina Cinquini, Davide BacciuICLR 2024 · 被引用 10 次
- Differentiable Structure Learning with Ancestral ConstraintsTaiyu Ban, Changxin Rong, Xiangyu Wang, Lyuzhou Chen 等ICML 2025
- Differentiable Structure Learning and Causal Discovery for General Binary DataChang Deng, Bryon AragamNeurIPS 2025
- Differentiable Structure Learning with Partial OrdersTaiyu Ban, Lyuzhou Chen, Xiangyu Wang, Xin Wang 等NeurIPS 2024 · 被引用 15 次
