Identifiable Nonlinear Differentiable Causal Discovery via Independence and Adaptive Group Sparsity
Ruicong Yao, Tim Verdonck, Mihaela van der Schaar, Jakob Raymaekers
Abstract
Differentiable approaches to causal discovery have shown promise in learning DAG structures via continuous optimization, but their theoretical guarantees are largely restricted to models with homoscedastic noise or known noise distribution. In particular, existing methods based on mean squared error fail to identify the true DAG when noise distributions are non-Gaussian and vary in scale. In this paper, we address this gap in nonlinear additive noise models (ANMs) with arbitrary noise. Our approach extends NOTIME (Berrevoets et al., 2025) which minimizes an independence criterion among the residuals. We show that the global minimizer of the independence criterion corresponds to the true underlying DAG up to additional constant edges in general ANMs. To recover the exact structure, we introduce an adaptive group lasso penalty that regularizes entire columns of the first-layer weight matrix of an MLP, enabling the selective pruning of constant edges in a functionally meaningful way. Empirically, our method achieves effective and stable performance across diverse noise types and variances, outperforming prior methods that lack identifiability guarantees in this setting.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on10
- Gradient-Based Neural DAG LearningSébastien Lachapelle, Philippe Brouillard, Tristan Deleu, Simon Lacoste-JulienICLR 2020 · 337 citations
- On the Role of Sparsity and DAG Constraints for Learning Linear DAGsIgnavier Ng, AmirEmad Ghassami, Kun ZhangNeurIPS 2020 · 306 citations
- Beware of the Simulated DAG! Causal Discovery Benchmarks May Be Easy to GameAlexander G. Reisach, Christof Seiler, Sebastian WeichwaldNeurIPS 2021 · 213 citations
- Score Matching Enables Causal Discovery of Nonlinear Additive Noise ModelsPaul Rolland, Volkan Cevher, Matthäus Kleindessner, Chris Russell et al.ICML 2022 · 123 citations
- DAGs with No Fears: A Closer Look at Continuous Optimization for Learning Bayesian NetworksDennis Wei, Tian Gao, Yue YuNeurIPS 2020 · 102 citations
Related papers
- DARING: Differentiable Causal Discovery with Residual IndependenceYue He, Peng Cui, Zheyan Shen, Renzhe Xu et al.KDD 2021 · 28 citations
- Stabilizing Causal Structure Learning under Heteroscedasticity: Analysis and Mitigation of Optimization FailuresEunjung Choi, Seonggyeom Kim, Dong-Kyu ChaeKDD 2026
- Boosting Causal Discovery via Adaptive Sample ReweightingAn Zhang, Fangfu Liu, Wenchang Ma, Zhibo Cai et al.ICLR 2023 · 4 citations
- Effective Causal Discovery under Identifiable Heteroscedastic Noise ModelNaiyu Yin, Tian Gao, Yue Yu, Qiang JiAAAI 2024 · 5 citations
- Strong and Weak Identifiability of Optimization-based Causal Discovery in Non-linear Additive Noise ModelsMingjia Li, Hong Qian, Tian-Zuo Wang, Shujun Li et al.ICML 2025
