A Scale-Invariant Sorting Criterion to Find a Causal Order in Additive Noise Models
Alexander G. Reisach, Myriam Tami, Christof Seiler, Antoine Chambaz, Sebastian Weichwald
摘要
Additive Noise Models (ANMs) are a common model class for causal discovery from observational data. Due to a lack of real-world data for which an underlying ANM is known, ANMs with randomly sampled parameters are commonly used to simulate data for the evaluation of causal discovery algorithms. While some parameters may be fixed by explicit assumptions, fully specifying an ANM requires choosing all parameters. Reisach et al. (2021) show that, for many ANM parameter choices, sorting the variables by increasing variance yields an ordering close to a causal order and introduce 'var-sortability' to quantify this alignment. Since increasing variances may be unrealistic and cannot be exploited when data scales are arbitrary, ANM data are often rescaled to unit variance in causal discovery benchmarking. We show that synthetic ANM data are characterized by another pattern that is scale-invariant and thus persists even after standardization: the explainable fraction of a variable's variance, as captured by the coefficient of determination R 2 , tends to increase along the causal order. The result is high 'R 2 -sortability', meaning that sorting the variables by increasing R 2 yields an ordering close to a causal order. We propose a computationally efficient baseline algorithm termed 'R 2 -SortnRegress' that exploits high R 2 -sortability and that can match and exceed the performance of established causal discovery algorithms. We show analytically that sufficiently high edge weights lead to a relative decrease of the noise contributions along causal chains, resulting in increasingly deterministic relationships and high R 2 . We characterize R 2 -sortability on synthetic data with different simulation parameters and find high values in common settings. Our findings reveal high R 2 -sortability as an assumption about the data generating process relevant to causal discovery and implicit in many ANM sampling schemes. It should be made explicit, as its prevalence in real-world data is an open question. For causal discovery benchmarking, we provide implementations of R 2 -sortability, the R 2 -SortnRegress algorithm, and ANM simulation procedures in our library CausalDisco.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- Root Cause Analysis of Outliers with Missing Structural KnowledgeWilliam Roy Orchard, Nastaran Okati, Sergio Hernan Garrido Mejia, Patrick Blöbaum 等NeurIPS 2025 · 被引用 24 次
- Embracing Discrete Search: A Reasonable Approach to Causal Structure LearningMarcel Wienöbst, Leonard Henckel, Sebastian WeichwaldICLR 2026 · 被引用 4 次
- Causal Discovery via Quantile Partial EffectYikang Chen, Xingzhe Sun, Dehui duICLR 2026 · 被引用 3 次
- Causal Mixture Models: Characterization and DiscoverySarah Mameche, Janis Kalofolias, Jilles VreekenNeurIPS 2025 · 被引用 1 次
- When Additive Noise Meets Unobserved Mediators: Bivariate Denoising Diffusion for Causal DiscoveryDominik Meier, Sujai Hiremath, Promit Ghosal, Kyra GanNeurIPS 2025 · 被引用 1 次
它引用的顶会 Paper6
- Beware of the Simulated DAG! Causal Discovery Benchmarks May Be Easy to GameAlexander G. Reisach, Christof Seiler, Sebastian WeichwaldNeurIPS 2021 · 被引用 213 次
- Score Matching Enables Causal Discovery of Nonlinear Additive Noise ModelsPaul Rolland, Volkan Cevher, Matthäus Kleindessner, Chris Russell 等ICML 2022 · 被引用 123 次
- Amortized Inference for Causal Structure LearningLars Lorch, Scott Sussex, Jonas Rothfuss, Andreas Krause 等NeurIPS 2022 · 被引用 118 次
- Assumption violations in causal discovery and the robustness of score matchingFrancesco Montagna, Atalanti-Anastasia Mastakouri, Elias Eulig, Nicoletta Noceti 等NeurIPS 2023 · 被引用 35 次
- Inferring Cause and Effect in the Presence of Heteroscedastic NoiseSascha Xu, Osman Mian, Alexander Marx, Jilles VreekenICML 2022 · 被引用 25 次
相关 Paper
- Standardizing Structural Causal ModelsWeronika Ormaniec, Scott Sussex, Lars Lorch, Bernhard Schölkopf 等ICLR 2025
- Strong and Weak Identifiability of Optimization-based Causal Discovery in Non-linear Additive Noise ModelsMingjia Li, Hong Qian, Tian-Zuo Wang, Shujun Li 等ICML 2025
- MissDAG: Causal Discovery in the Presence of Missing Data with Continuous Additive Noise ModelsErdun Gao, Ignavier Ng, Mingming Gong, Li Shen 等NeurIPS 2022 · 被引用 36 次
- Score-informed Neural Operator for Enhancing Ordering-based Causal DiscoveryJiyeon Kang, Songseong Kim, Chanhui Lee, Doyeong Hwang 等NeurIPS 2025 · 被引用 2 次
- Identification of Causal Structure in the Presence of Missing Data with Additive Noise ModelJie Qiao, Zhengming Chen, Jianhua Yu, Ruichu Cai 等AAAI 2024 · 被引用 7 次
