Beware of the Simulated DAG! Causal Discovery Benchmarks May Be Easy to Game
Alexander G. Reisach, Christof Seiler, Sebastian Weichwald
Abstract
Simulated DAG models may exhibit properties that, perhaps inadvertently, render their structure identifiable and unexpectedly affect structure learning algorithms. Here, we show that marginal variance tends to increase along the causal order for generically sampled additive noise models. We introduce varsortability as a measure of the agreement between the order of increasing marginal variance and the causal order. For commonly sampled graphs and model parameters, we show that the remarkable performance of some continuous structure learning algorithms can be explained by high varsortability and matched by a simple baseline method. Yet, this performance may not transfer to real-world data where varsortability may be moderate or dependent on the choice of measurement scales. On standardized data, the same algorithms fail to identify the ground-truth DAG or its Markov equivalence class. While standardization removes the pattern in marginal variance, we show that data generating processes that incur high varsortability also leave a distinct covariance pattern that may be exploited even after standardization. Our findings challenge the significance of generic benchmarks with independently drawn parameters. The code is available at https://github.com/Scriddie/ Varsortability .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext f07416dd-2cb8-4c96-a0c6-fb42d00b8e77Cited by top-tier papers58
- Learning Linear Causal Representations from Interventions under General Nonlinear MixingSimon Buchholz, Goutham Rajendran, Elan Rosenfeld, Bryon Aragam et al.NeurIPS 2023 · 113 citations
- Large-Scale Differentiable Causal Discovery of Factor GraphsRomain Lopez, Jan-Christian Hütter, Jonathan K. Pritchard, Aviv RegevNeurIPS 2022 · 78 citations
- On the Identifiability and Estimation of Causal Location-Scale Noise ModelsAlexander Immer, Christoph Schultheiss, Julia E. Vogt, Bernhard Schölkopf et al.ICML 2023 · 56 citations
- A Scale-Invariant Sorting Criterion to Find a Causal Order in Additive Noise ModelsAlexander G. Reisach, Myriam Tami, Christof Seiler, Antoine Chambaz et al.NeurIPS 2023 · 40 citations
- CausalTime: Realistically Generated Time-series for Benchmarking of Causal DiscoveryYuxiao Cheng, Ziqian Wang, Tingxiong Xiao, Qin Zhong et al.ICLR 2024 · 35 citations
Builds on5
- Gradient-Based Neural DAG LearningSébastien Lachapelle, Philippe Brouillard, Tristan Deleu, Simon Lacoste-JulienICLR 2020 · 337 citations
- On the Role of Sparsity and DAG Constraints for Learning Linear DAGsIgnavier Ng, AmirEmad Ghassami, Kun ZhangNeurIPS 2020 · 306 citations
- Differentiable Causal Discovery from Interventional DataPhilippe Brouillard, Sébastien Lachapelle, Alexandre Lacoste, Simon Lacoste-Julien et al.NeurIPS 2020 · 295 citations
- DAGs with No Fears: A Closer Look at Continuous Optimization for Learning Bayesian NetworksDennis Wei, Tian Gao, Yue YuNeurIPS 2020 · 102 citations
- A polynomial-time algorithm for learning nonparametric causal graphsMing Gao, Yi Ding, Bryon AragamNeurIPS 2020 · 39 citations
Related papers
- Standardizing Structural Causal ModelsWeronika Ormaniec, Scott Sussex, Lars Lorch, Bernhard Schölkopf et al.ICLR 2025
- Stabilizing Causal Structure Learning under Heteroscedasticity: Analysis and Mitigation of Optimization FailuresEunjung Choi, Seonggyeom Kim, Dong-Kyu ChaeKDD 2026
- Assumption violations in causal discovery and the robustness of score matchingFrancesco Montagna, Atalanti-Anastasia Mastakouri, Elias Eulig, Nicoletta Noceti et al.NeurIPS 2023 · 35 citations
- Identification of Linear Latent Variable Model with Arbitrary DistributionZhengming Chen, Feng Xie, Jie Qiao, Zhifeng Hao et al.AAAI 2022 · 24 citations
- Deriving Causal Order from Single-Variable Interventions: Guarantees & AlgorithmMathieu Chevalley, Patrick Schwab, Arash MehrjouICLR 2025
