Smooth Multi-Policy Causal Effect Estimation in Longitudinal Settings
Wenxin Chen, Weishen Pan, Kyra Gan, Fei Wang
Abstract
Comparative evaluation of multiple dynamic treatment policies is essential for healthcare and policy decisions, yet conventional longitudinal causal inference methods estimate each in isolation , preventing information sharing across counterfactuals. We demonstrate that this separate estimation paradigm induces a structurally uncontrolled second-order bias, inflating finite-sample variance even after standard debiasing with longitudinal targeted maximum likelihood estimation (LTMLE). To address this, we propose a policy-aware reparameterization of Iterative Conditional Expectation (ICE) Q-functions that enables joint estimation through shared representations. We implement this approach in the Policy-Encoded Q Network (PEQ-Net) , an architecture centered on a shared policy encoder. The encoder is trained using kernel mean embeddings, ensuring that the learned representation space reflects population-level policy dissimilarities. After applying an LTMLE correction step, we prove this design imposes a structural constraint on the second-order remainder, thereby stabilizing finite-sample variance. Experiments on semi-synthetic datasets demonstrate that PEQ-Net consistently outperforms existing ICE-based methods, achieving substantial reductions in root-mean-square error, particularly when evaluating closely related policies.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 1769caee-40e7-43e5-897c-2f5021230fbcBuilds on14
- Estimating counterfactual treatment outcomes over time through adversarially balanced representationsIoana Bica, Ahmed M. Alaa, James Jordon, Mihaela van der SchaarICLR 2020 · 224 citations
- Minimax-Optimal Off-Policy Evaluation with Linear Function ApproximationYaqi Duan, Zeyu Jia, Mengdi WangICML 2020 · 161 citations
- Causal Transformer for Estimating Counterfactual OutcomesValentyn Melnychuk, Dennis Frauen, Stefan FeuerriegelICML 2022 · 146 citations
- Continuous-Time Modeling of Counterfactual Outcomes Using Neural Controlled Differential EquationsNabeel Seedat, Fergus Imrie, Alexis Bellot, Zhaozhi Qian et al.ICML 2022 · 68 citations
- Estimating Average Causal Effects from Patient TrajectoriesDennis Frauen, Tobias Hatt, Valentyn Melnychuk, Stefan FeuerriegelAAAI 2023 · 34 citations
Related papers
- Longitudinal Targeted Minimum Loss-based Estimation with Temporal-Difference Heterogeneous TransformerToru Shirakawa, Yi Li, Yulun Wu, Sky Qiu et al.ICML 2024 · 18 citations
- Invariant Causal Imitation Learning for Generalizable PoliciesIoana Bica, Daniel Jarrett, Mihaela van der SchaarNeurIPS 2021 · 46 citations
- Debiased Model-based Representations for Sample-efficient Continuous ControlJiafei Lyu, Zichuan Lin, Scott Fujimoto, Kai Yang et al.ICML 2026
- On Inductive Biases for Heterogeneous Treatment Effect EstimationAlicia Curth, Mihaela van der SchaarNeurIPS 2021 · 114 citations
- Iterative Amortized Policy OptimizationJoseph Marino, Alexandre Piché, Alessandro Davide Ialongo, Yisong YueNeurIPS 2021 · 27 citations
