Causal de Finetti: On the Identification of Invariant Causal Structure in Exchangeable Data
Siyuan Guo, Viktor Tóth, Bernhard Schölkopf, Ferenc Huszar
Abstract
Constraint-based causal discovery methods leverage conditional independence tests to infer causal relationships in a wide variety of applications. Just as the majority of machine learning methods, existing work focuses on studying independent and identically distributed data. However, it is known that even with infinite i.i.d. data, constraint-based methods can only identify causal structures up to broad Markov equivalence classes, posing a fundamental limitation for causal discovery. In this work, we observe that exchangeable data contains richer conditional independence structure than i.i.d. data, and show how the richer structure can be leveraged for causal discovery. We first present causal de Finetti theorems, which state that exchangeable distributions with certain non-trivial conditional independences can always be represented as independent causal mechanism (ICM) generative processes. We then present our main identifiability theorem, which shows that given data from an ICM generative process, its unique causal structure can be identified through performing conditional independence tests. We finally develop a causal discovery algorithm and demonstrate its applicability to inferring causal relationships from multi-environment data. Our code and models are publicly available at: https://github.com/syguo96/Causal-de-Finetti * Equal contribution 37th Conference on Neural Information Processing Systems (NeurIPS 2023).
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 1d732e35-46cc-4ced-a02f-2af9ee58cfabCited by top-tier papers16
- Causal Discovery in Heterogeneous Environments Under the Sparse Mechanism Shift HypothesisRonan Perry, Julius von Kügelgen, Bernhard SchölkopfNeurIPS 2022 · 84 citations
- Object Representations as Fixed Points: Training Iterative Refinement Algorithms with Implicit DifferentiationMichael Chang, Tom Griffiths, Sergey LevineNeurIPS 2022 · 69 citations
- Do-PFN: In-Context Learning for Causal Effect EstimationJake Robertson, Arik Reuter, Siyuan Guo, Noah Hollmann et al.NeurIPS 2025 · 58 citations
- Detecting hidden confounding in observational data using multiple environmentsRickard Karlsson, Jesse H. KrijtheNeurIPS 2023 · 22 citations
- Learning Causal Models under Independent ChangesSarah Mameche, David Kaltenpoth, Jilles VreekenNeurIPS 2023 · 15 citations
Builds on4
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Recurrent Independent MechanismsAnirudh Goyal, Alex Lamb, Jordan Hoffmann, Shagun Sodhani et al.ICLR 2021 · 357 citations
- Weakly supervised causal representation learningJohann Brehmer, Pim de Haan, Phillip Lippe, Taco S. CohenNeurIPS 2022 · 196 citations
- Fast And Slow Learning Of Recurrent Independent MechanismsKanika Madan, Nan Rosemary Ke, Anirudh Goyal, Bernhard Schölkopf et al.ICLR 2021 · 41 citations
Related papers
- Identifiable Exchangeable Mechanisms for Causal Structure and Representation LearningPatrik Reizinger, Siyuan Guo, Ferenc Huszár, Bernhard Schölkopf et al.ICLR 2025
- Do Finetti: On Causal Effects for Exchangeable DataSiyuan Guo, Chi Zhang, Karthika Mohan, Ferenc Huszar et al.NeurIPS 2024 · 12 citations
- Integrating Overlapping Datasets Using Bivariate Causal DiscoveryAnish Dhir, Ciarán M. LeeAAAI 2020 · 23 citations
- Characterization and Learning of Causal Graphs with Small Conditioning SetsMurat KocaogluNeurIPS 2023 · 17 citations
- On the identifiability of causal graphs with multiple environmentsFrancesco MontagnaICLR 2026 · 2 citations
