Sample Complexity of Interventional Causal Representation Learning
Emre Acartürk, Burak Varici, Karthikeyan Shanmugam, Ali Tajer
Abstract
Consider a data-generation process that transforms low-dimensional latent causally-related variables to high-dimensional observed variables. Causal representation learning (CRL) is the process of using the observed data to recover the latent causal variables and the causal structure among them. Despite the multitude of identifiability results under various interventional CRL settings, the existing guarantees apply exclusively to the infinite-sample regime (i.e., infinite observed samples). This paper establishes the first sample-complexity analysis for the finite-sample regime, in which the interactions between the number of observed samples and probabilistic guarantees on recovering the latent variables and structure are established. This paper focuses on general latent causal models, stochastic soft interventions, and a linear transformation from the latent to the observation space. The identifiability results ensure graph recovery up to ancestors and latent variables recovery up to mixing with parent variables. Specifically, O ((log 1 δ ) 4 ) samples suffice for latent graph recovery up to ancestors with probability 1 − δ , and O (( 1 ϵ log 1 δ ) 4 ) samples suffice for latent causal variables recovery that is ϵ close to the identifiability class with probability 1 − δ .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b06f6d36-1b58-446c-ab2f-a885b81ca515Cited by top-tier papers3
- Sample-efficient Learning of Concepts with Theoretical Guarantees: from Data to Concepts without InterventionsHidde Fokkema, Tim van Erven, Sara MagliacaneNeurIPS 2025 · 7 citations
- Linear Causal Representation Learning by Topological Ordering, Pruning, and DisentanglementHao Chen, Lin Liu, Yuguang WangICML 2026
- Reward-oriented Causal Representation LearningZirui Yan, Emre Acartürk, Ali TajerNeurIPS 2025
Builds on7
- Interventional Causal Representation LearningKartik Ahuja, Divyat Mahajan, Yixin Wang, Yoshua BengioICML 2023 · 143 citations
- Nonparametric Identifiability of Causal Representations from Unknown InterventionsJulius von Kügelgen, Michel Besserve, Wendong Liang, Luigi Gresele et al.NeurIPS 2023 · 127 citations
- Identifiability Guarantees for Causal Disentanglement from Soft InterventionsJiaqi Zhang, Kristjan H. Greenewald, Chandler Squires, Akash Srivastava et al.NeurIPS 2023 · 120 citations
- Learning Linear Causal Representations from Interventions under General Nonlinear MixingSimon Buchholz, Goutham Rajendran, Elan Rosenfeld, Bryon Aragam et al.NeurIPS 2023 · 113 citations
- Linear Causal Disentanglement via InterventionsChandler Squires, Anna Seigal, Salil S. Bhate, Caroline UhlerICML 2023 · 90 citations
Related papers
- Linear Causal Representation Learning from Unknown Multi-node InterventionsBurak Varici, Emre Acartürk, Karthikeyan Shanmugam, Ali TajerNeurIPS 2024 · 19 citations
- Learning Linear Causal Representations from General Environments: Identifiability and Intrinsic AmbiguityJikai Jin, Vasilis SyrgkanisNeurIPS 2024 · 10 citations
- Causal Representation Learning Made Identifiable by Grouping of Observational VariablesHiroshi Morioka, Aapo HyvärinenICML 2024 · 26 citations
- Causal Component AnalysisWendong Liang, Armin Kekic, Julius von Kügelgen, Simon Buchholz et al.NeurIPS 2023 · 65 citations
- Causal Representation Learning from Multiple Distributions: A General SettingKun Zhang, Shaoan Xie, Ignavier Ng, Yujia ZhengICML 2024 · 61 citations
