Beware of the Simulated DAG! Causal Discovery Benchmarks May Be Easy to Game
Alexander G. Reisach, Christof Seiler, Sebastian Weichwald
摘要
Simulated DAG models may exhibit properties that, perhaps inadvertently, render their structure identifiable and unexpectedly affect structure learning algorithms. Here, we show that marginal variance tends to increase along the causal order for generically sampled additive noise models. We introduce varsortability as a measure of the agreement between the order of increasing marginal variance and the causal order. For commonly sampled graphs and model parameters, we show that the remarkable performance of some continuous structure learning algorithms can be explained by high varsortability and matched by a simple baseline method. Yet, this performance may not transfer to real-world data where varsortability may be moderate or dependent on the choice of measurement scales. On standardized data, the same algorithms fail to identify the ground-truth DAG or its Markov equivalence class. While standardization removes the pattern in marginal variance, we show that data generating processes that incur high varsortability also leave a distinct covariance pattern that may be exploited even after standardization. Our findings challenge the significance of generic benchmarks with independently drawn parameters. The code is available at https://github.com/Scriddie/ Varsortability .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper58
- Learning Linear Causal Representations from Interventions under General Nonlinear MixingSimon Buchholz, Goutham Rajendran, Elan Rosenfeld, Bryon Aragam 等NeurIPS 2023 · 被引用 113 次
- Large-Scale Differentiable Causal Discovery of Factor GraphsRomain Lopez, Jan-Christian Hütter, Jonathan K. Pritchard, Aviv RegevNeurIPS 2022 · 被引用 78 次
- On the Identifiability and Estimation of Causal Location-Scale Noise ModelsAlexander Immer, Christoph Schultheiss, Julia E. Vogt, Bernhard Schölkopf 等ICML 2023 · 被引用 56 次
- A Scale-Invariant Sorting Criterion to Find a Causal Order in Additive Noise ModelsAlexander G. Reisach, Myriam Tami, Christof Seiler, Antoine Chambaz 等NeurIPS 2023 · 被引用 40 次
- CausalTime: Realistically Generated Time-series for Benchmarking of Causal DiscoveryYuxiao Cheng, Ziqian Wang, Tingxiong Xiao, Qin Zhong 等ICLR 2024 · 被引用 35 次
它引用的顶会 Paper5
- Gradient-Based Neural DAG LearningSébastien Lachapelle, Philippe Brouillard, Tristan Deleu, Simon Lacoste-JulienICLR 2020 · 被引用 337 次
- On the Role of Sparsity and DAG Constraints for Learning Linear DAGsIgnavier Ng, AmirEmad Ghassami, Kun ZhangNeurIPS 2020 · 被引用 306 次
- Differentiable Causal Discovery from Interventional DataPhilippe Brouillard, Sébastien Lachapelle, Alexandre Lacoste, Simon Lacoste-Julien 等NeurIPS 2020 · 被引用 295 次
- DAGs with No Fears: A Closer Look at Continuous Optimization for Learning Bayesian NetworksDennis Wei, Tian Gao, Yue YuNeurIPS 2020 · 被引用 102 次
- A polynomial-time algorithm for learning nonparametric causal graphsMing Gao, Yi Ding, Bryon AragamNeurIPS 2020 · 被引用 39 次
相关 Paper
- Standardizing Structural Causal ModelsWeronika Ormaniec, Scott Sussex, Lars Lorch, Bernhard Schölkopf 等ICLR 2025
- Stabilizing Causal Structure Learning under Heteroscedasticity: Analysis and Mitigation of Optimization FailuresEunjung Choi, Seonggyeom Kim, Dong-Kyu ChaeKDD 2026
- Assumption violations in causal discovery and the robustness of score matchingFrancesco Montagna, Atalanti-Anastasia Mastakouri, Elias Eulig, Nicoletta Noceti 等NeurIPS 2023 · 被引用 35 次
- Identification of Linear Latent Variable Model with Arbitrary DistributionZhengming Chen, Feng Xie, Jie Qiao, Zhifeng Hao 等AAAI 2022 · 被引用 24 次
- Deriving Causal Order from Single-Variable Interventions: Guarantees & AlgorithmMathieu Chevalley, Patrick Schwab, Arash MehrjouICLR 2025
