Amortized Active Causal Induction with Deep Reinforcement Learning
Yashas Annadani, Panagiotis Tigas, Stefan Bauer, Adam Foster
Abstract
We present Causal Amortized Active Structure Learning (CAASL), an active intervention design policy that can select interventions that are adaptive, real-time and that does not require access to the likelihood. This policy, an amortized network based on the transformer, is trained with reinforcement learning on a simulator of the design environment, and a reward function that measures how close the true causal graph is to a causal graph posterior inferred from the gathered data. On synthetic data and a single-cell gene expression simulator, we demonstrate empirically that the data acquired through our policy results in a better estimate of the underlying causal graph than alternative strategies. Our design policy successfully achieves amortized intervention design on the distribution of the training environment while also generalizing well to distribution shifts in test-time design environments. Further, our policy also demonstrates excellent zero-shot generalization to design environments with dimensionality higher than that during training, and to intervention types that it has not been trained on.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 688a31d3-de59-4185-805b-ce6cf16cca0dCited by top-tier papers3
- ActiveCQ: Active Estimation of Causal QuantitiesErdun Gao, Dino SejdinovicICLR 2026 · 1 citation
- Designing Time Series Experiments in A/B Testing with Transformer Reinforcement LearningXiangkun Wu, Qianglin Wen, Yingying Zhang, Hongtu Zhu et al.ICLR 2026 · 1 citation
- Causal Preference ElicitationEdwin V. Bonilla, He Zhao, Daniel M SteinbergICML 2026
Builds on21
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Scaling Vision TransformersXiaohua Zhai, Alexander Kolesnikov, Neil Houlsby, Lucas BeyerCVPR 2022 · 767 citations
- Differentiable Causal Discovery from Interventional DataPhilippe Brouillard, Sébastien Lachapelle, Alexandre Lacoste, Simon Lacoste-Julien et al.NeurIPS 2020 · 295 citations
- Causal Discovery with Reinforcement LearningShengyu Zhu, Ignavier Ng, Zhitang ChenICLR 2020 · 285 citations
- Beware of the Simulated DAG! Causal Discovery Benchmarks May Be Easy to GameAlexander G. Reisach, Christof Seiler, Sebastian WeichwaldNeurIPS 2021 · 213 citations
Related papers
- Amortized Inference for Causal Structure LearningLars Lorch, Scott Sussex, Jonas Rothfuss, Andreas Krause et al.NeurIPS 2022 · 118 citations
- Interventions, Where and How? Experimental Design for Causal Models at ScalePanagiotis Tigas, Yashas Annadani, Andrew Jesson, Bernhard Schölkopf et al.NeurIPS 2022 · 68 citations
- Learning to Induce Causal StructureNan Rosemary Ke, Silvia Chiappa, Jane X. Wang, Jörg Bornschein et al.ICLR 2023 · 17 citations
- Reinforcement Learning of Causal Variables Using Mediation AnalysisTue Herlau, Rasmus LarsenAAAI 2022 · 8 citations
- Near-Optimal Experiment Design in Linear non-Gaussian Cyclic ModelsEhsan Sharifian, Saber Salehkaleybar, Negar KiyavashNeurIPS 2025 · 4 citations
