Finding Counterfactually Optimal Action Sequences in Continuous State Spaces
Stratis Tsirtsis, Manuel Gomez Rodriguez
摘要
Whenever a clinician reflects on the efficacy of a sequence of treatment decisions for a patient, they may try to identify critical time steps where, had they made different decisions, the patient's health would have improved. While recent methods at the intersection of causal inference and reinforcement learning promise to aid human experts, as the clinician above, to retrospectively analyze sequential decision making processes, they have focused on environments with finitely many discrete states. However, in many practical applications, the state of the environment is inherently continuous in nature. In this paper, we aim to fill this gap. We start by formally characterizing a sequence of discrete actions and continuous states using finite horizon Markov decision processes and a broad class of bijective structural causal models. Building upon this characterization, we formalize the problem of finding counterfactually optimal action sequences and show that, in general, we cannot expect to solve it in polynomial time. Then, we develop a search method based on the A * algorithm that, under a natural form of Lipschitz continuity of the environment's dynamics, is guaranteed to return the optimal solution to the problem. Experiments on real clinical data show that our method is very efficient in practice, and it has the potential to offer interesting insights for sequential decision making tasks.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Learning Counterfactual Outcomes Under Rank PreservationPeng Wu, Haoxuan Li, Chunyuan Zheng, Yan Zeng 等NeurIPS 2025 · 被引用 7 次
- Counterfactual Identifiability via Dynamic Optimal TransportFabio De Sousa Ribeiro, Ainkaran Santhirasekaram, Ben GlockerNeurIPS 2025 · 被引用 7 次
- Exogenous Matching: Learning Good Proposals for Tractable Counterfactual EstimationYikang Chen, Dehui Du, Lili TianNeurIPS 2024 · 被引用 3 次
- Exogenous Isomorphism for Counterfactual IdentifiabilityYikang Chen, Dehui DuICML 2025
- Variational Counterfactual Intervention Planning to Achieve Target OutcomesXin Wang, Shengfei Lyu, Chi Luo, Xiren Zhou 等ICML 2025
它引用的顶会 Paper22
- Model Based Reinforcement Learning for AtariLukasz Kaiser, Mohammad Babaeizadeh, Piotr Milos, Blazej Osinski 等ICLR 2020 · 被引用 969 次
- Explainable Reinforcement Learning through a Causal LensPrashan Madumal, Tim Miller, Liz Sonenberg, Frank VetereAAAI 2020 · 被引用 408 次
- Deep Structural Causal Models for Tractable Counterfactual InferenceNick Pawlowski, Daniel Coelho de Castro, Ben GlockerNeurIPS 2020 · 被引用 353 次
- Algorithmic recourse under imperfect causal knowledge: a probabilistic approachAmir-Hossein Karimi, Bodo Julius von Kügelgen, Bernhard Schölkopf, Isabel ValeraNeurIPS 2020 · 被引用 224 次
- BCD Nets: Scalable Variational Approaches for Bayesian Causal DiscoveryChris Cundy, Aditya Grover, Stefano ErmonNeurIPS 2021 · 被引用 105 次
相关 Paper
- Counterfactual Explanations in Sequential Decision Making Under UncertaintyStratis Tsirtsis, Abir De, Manuel Gomez RodriguezNeurIPS 2021 · 被引用 59 次
- Causal Modeling of Policy Interventions From Treatment-Outcome SequencesCaglar Hizli, S. T. John, Anne Tuulikki Juuti, Tuure Tapani Saarinen 等ICML 2023 · 被引用 7 次
- Learning to search efficiently for causally near-optimal treatmentsSamuel Håkansson, Viktor Lindblom, Omer Gottesman, Fredrik D. JohanssonNeurIPS 2020 · 被引用 7 次
- Dynamic Causal Bayesian OptimizationVirginia Aglietti, Neil Dhir, Javier González, Theodoros DamoulasNeurIPS 2021 · 被引用 40 次
- Exploiting Causal Graph Priors with Posterior Sampling for Reinforcement LearningMirco Mutti, Riccardo De Santi, Marcello Restelli, Alexander Marx 等ICLR 2024 · 被引用 6 次
