Finding Counterfactually Optimal Action Sequences in Continuous State Spaces
Stratis Tsirtsis, Manuel Gomez Rodriguez
Abstract
Whenever a clinician reflects on the efficacy of a sequence of treatment decisions for a patient, they may try to identify critical time steps where, had they made different decisions, the patient's health would have improved. While recent methods at the intersection of causal inference and reinforcement learning promise to aid human experts, as the clinician above, to retrospectively analyze sequential decision making processes, they have focused on environments with finitely many discrete states. However, in many practical applications, the state of the environment is inherently continuous in nature. In this paper, we aim to fill this gap. We start by formally characterizing a sequence of discrete actions and continuous states using finite horizon Markov decision processes and a broad class of bijective structural causal models. Building upon this characterization, we formalize the problem of finding counterfactually optimal action sequences and show that, in general, we cannot expect to solve it in polynomial time. Then, we develop a search method based on the A * algorithm that, under a natural form of Lipschitz continuity of the environment's dynamics, is guaranteed to return the optimal solution to the problem. Experiments on real clinical data show that our method is very efficient in practice, and it has the potential to offer interesting insights for sequential decision making tasks.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 8f5030a0-c126-4f19-bfbe-e942790bc1dfCited by top-tier papers5
- Learning Counterfactual Outcomes Under Rank PreservationPeng Wu, Haoxuan Li, Chunyuan Zheng, Yan Zeng et al.NeurIPS 2025 · 7 citations
- Counterfactual Identifiability via Dynamic Optimal TransportFabio De Sousa Ribeiro, Ainkaran Santhirasekaram, Ben GlockerNeurIPS 2025 · 7 citations
- Exogenous Matching: Learning Good Proposals for Tractable Counterfactual EstimationYikang Chen, Dehui Du, Lili TianNeurIPS 2024 · 3 citations
- Exogenous Isomorphism for Counterfactual IdentifiabilityYikang Chen, Dehui DuICML 2025
- Variational Counterfactual Intervention Planning to Achieve Target OutcomesXin Wang, Shengfei Lyu, Chi Luo, Xiren Zhou et al.ICML 2025
Builds on22
- Model Based Reinforcement Learning for AtariLukasz Kaiser, Mohammad Babaeizadeh, Piotr Milos, Blazej Osinski et al.ICLR 2020 · 969 citations
- Explainable Reinforcement Learning through a Causal LensPrashan Madumal, Tim Miller, Liz Sonenberg, Frank VetereAAAI 2020 · 408 citations
- Deep Structural Causal Models for Tractable Counterfactual InferenceNick Pawlowski, Daniel Coelho de Castro, Ben GlockerNeurIPS 2020 · 353 citations
- Algorithmic recourse under imperfect causal knowledge: a probabilistic approachAmir-Hossein Karimi, Bodo Julius von Kügelgen, Bernhard Schölkopf, Isabel ValeraNeurIPS 2020 · 224 citations
- BCD Nets: Scalable Variational Approaches for Bayesian Causal DiscoveryChris Cundy, Aditya Grover, Stefano ErmonNeurIPS 2021 · 105 citations
Related papers
- Counterfactual Explanations in Sequential Decision Making Under UncertaintyStratis Tsirtsis, Abir De, Manuel Gomez RodriguezNeurIPS 2021 · 59 citations
- Causal Modeling of Policy Interventions From Treatment-Outcome SequencesCaglar Hizli, S. T. John, Anne Tuulikki Juuti, Tuure Tapani Saarinen et al.ICML 2023 · 7 citations
- Learning to search efficiently for causally near-optimal treatmentsSamuel Håkansson, Viktor Lindblom, Omer Gottesman, Fredrik D. JohanssonNeurIPS 2020 · 7 citations
- Dynamic Causal Bayesian OptimizationVirginia Aglietti, Neil Dhir, Javier González, Theodoros DamoulasNeurIPS 2021 · 40 citations
- Exploiting Causal Graph Priors with Posterior Sampling for Reinforcement LearningMirco Mutti, Riccardo De Santi, Marcello Restelli, Alexander Marx et al.ICLR 2024 · 6 citations
