Learning Generalized Gumbel-max Causal Mechanisms
Guy Lorberbom, Daniel D. Johnson, Chris J. Maddison, Daniel Tarlow, Tamir Hazan
Abstract
To perform counterfactual reasoning in Structural Causal Models (SCMs), one needs to know the causal mechanisms, which provide factorizations of conditional distributions into noise sources and deterministic functions mapping realizations of noise to samples. Unfortunately, the causal mechanism is not uniquely identified by data that can be gathered by observing and interacting with the world, so there remains the question of how to choose causal mechanisms. In recent work, Oberst & Sontag (2019) propose Gumbel-max SCMs, which use Gumbel-max reparameterizations as the causal mechanism due to an intuitively appealing counterfactual stability property. In this work, we instead argue for choosing a causal mechanism that is best under a quantitative criteria such as minimizing variance when estimating counterfactual treatment effects. We propose a parameterized family of causal mechanisms that generalize Gumbel-max. We show that they can be trained to minimize counterfactual effect variance and other losses on a distribution of queries of interest, yielding lower variance estimates of counterfactual treatment effect than fixed alternatives, also generalizing to queries not seen at training time.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0ff12963-9708-463b-afc6-2d5b515d088dCited by top-tier papers5
- Counterfactual Identifiability of Bijective Causal ModelsArash Nasr-Esfahany, Mohammad Alizadeh, Devavrat ShahICML 2023 · 42 citations
- Deep Counterfactual Estimation with Categorical Background VariablesEdward De BrouwerNeurIPS 2022 · 8 citations
- Counterfactual Analysis in Dynamic Latent State ModelsMartin B. Haugh, Raghav SingalICML 2023 · 8 citations
- Curiosity in Hindsight: Intrinsic Exploration in Stochastic EnvironmentsDaniel Jarrett, Corentin Tallec, Florent Altché, Thomas Mesnard et al.ICML 2023 · 3 citations
- Off-Policy Selection for Initiating Human-Centric Experimental DesignGe Gao, Xi Yang, Qitong Gao, Song Ju et al.NeurIPS 2024 · 1 citation
Related papers
- Causal Discovery and Inference through Next-Token PredictionEivinas Butkus, Nikolaus KriegeskorteNeurIPS 2025 · 3 citations
- Gumbel Counterfactual Generation From Language ModelsShauli Ravfogel, Anej Svete, Vésteinn Snæbjarnarson, Ryan CotterellICLR 2025
- Answering Complex Causal Queries With the Maximum Causal Set EffectZachary MarkovichNeurIPS 2021
- High Fidelity Image Counterfactuals with Probabilistic Causal ModelsFabio De Sousa Ribeiro, Tian Xia, Miguel Monteiro, Nick Pawlowski et al.ICML 2023 · 68 citations
- Language Models as Causal Effect GeneratorsLucius E. J. Bynum, Kyunghyun ChoEMNLP 2025
