Truncated Matrix Power Iteration for Differentiable DAG Learning
Zhen Zhang, Ignavier Ng, Dong Gong, Yuhang Liu, Ehsan Abbasnejad, Mingming Gong, Kun Zhang, Javen Qinfeng Shi
摘要
Recovering underlying Directed Acyclic Graph (DAG) structures from observational data is highly challenging due to the combinatorial nature of the DAGconstrained optimization problem. Recently, DAG learning has been cast as a continuous optimization problem by characterizing the DAG constraint as a smooth equality one, generally based on polynomials over adjacency matrices. Existing methods place very small coefficients on high-order polynomial terms for stabilization, since they argue that large coefficients on the higher-order terms are harmful due to numeric exploding. On the contrary, we discover that large coefficients on higher-order terms are beneficial for DAG learning, when the spectral radiuses of the adjacency matrices are small, and that larger coefficients for higher-order terms can approximate the DAG constraints much better than the small counterparts. Based on this, we propose a novel DAG learning method with efficient truncated matrix power iteration to approximate geometric series based DAG constraints. Empirically, our DAG learning method outperforms the previous state-of-the-arts in various settings, often by a factor of 3 or more in terms of structural Hamming distance. * Equal Contribution Preprint. Under review.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper12
- Constraint-Free Structure Learning with Smooth Acyclic OrientationsRiccardo Massidda, Francesco Landolfi, Martina Cinquini, Davide BacciuICLR 2024 · 被引用 10 次
- Learning Large DAGs is Harder than you Think: Many Losses are Minimal for the Wrong DAGJonas Seng, Matej Zecevic, Devendra Singh Dhami, Kristian KerstingICLR 2024 · 被引用 9 次
- On the Identifiability of Sparse ICA without Assuming Non-GaussianityIgnavier Ng, Yujia Zheng, Xinshuai Dong, Kun ZhangNeurIPS 2023 · 被引用 9 次
- Markov Equivalence and Consistency in Differentiable Structure LearningChang Deng, Kevin Bello, Pradeep Ravikumar, Bryon AragamNeurIPS 2024 · 被引用 8 次
- Energy Efficient Streaming Time Series Classification with Attentive Power IterationHao Huang, Tapan Shah, Scott Evans, Shinjae YooAAAI 2024 · 被引用 4 次
它引用的顶会 Paper5
- Gradient-Based Neural DAG LearningSébastien Lachapelle, Philippe Brouillard, Tristan Deleu, Simon Lacoste-JulienICLR 2020 · 被引用 337 次
- On the Role of Sparsity and DAG Constraints for Learning Linear DAGsIgnavier Ng, AmirEmad Ghassami, Kun ZhangNeurIPS 2020 · 被引用 306 次
- Causal Discovery with Reinforcement LearningShengyu Zhu, Ignavier Ng, Zhitang ChenICLR 2020 · 被引用 285 次
- DAGs with No Fears: A Closer Look at Continuous Optimization for Learning Bayesian NetworksDennis Wei, Tian Gao, Yue YuNeurIPS 2020 · 被引用 102 次
- DAGs with No Curl: An Efficient DAG Structure Learning ApproachYue Yu, Tian Gao, Naiyu Yin, Qiang JiICML 2021 · 被引用 77 次
相关 Paper
- Analytic DAG Constraints for Differentiable DAG LearningZhen Zhang, Ignavier Ng, Dong Gong, Yuhang Liu 等ICLR 2025
- ProDAG: Projected Variational Inference for Directed Acyclic GraphsRyan Thompson, Edwin V. Bonilla, Robert KohnNeurIPS 2025 · 被引用 6 次
- CoLiDE: Concomitant Linear DAG EstimationSeyed Saman Saboksayr, Gonzalo Mateos, Mariano TepperICLR 2024 · 被引用 9 次
- Causal Discovery via Bayesian OptimizationBao Duong, Sunil Gupta, Thin NguyenICLR 2025
- DAG Learning on the PermutahedronValentina Zantedeschi, Luca Franceschi, Jean Kaddour, Matt J. Kusner 等ICLR 2023
