A Two-Stage Pretraining-Finetuning Framework for Treatment Effect Estimation with Unmeasured Confounding
Chuan Zhou, Yaxuan Li, Chunyuan Zheng, Haiteng Zhang, Min Zhang, Haoxuan Li, Mingming Gong
摘要
Estimating the conditional average treatment effect (CATE) from observational data plays a crucial role in areas such as e-commerce, healthcare, and economics. Existing studies mainly rely on the strong ignorability assumption that there are no unmeasured confounders, whose presence cannot be tested from observational data and can invalidate any causal conclusion. In contrast, data collected from randomized controlled trials (RCT) do not suffer from confounding, but are usually limited by a small sample size. In this paper, we propose a two-stage pretraining-finetuning (TSPF) framework using both large-scale observational data and small-scale RCT data to estimate the CATE in the presence of unmeasured confounding. In the first stage, a foundational representation of covariates is trained to estimate counterfactual outcomes through large-scale observational data. In the second stage, we propose to train an augmented representation of the covariates, which is concatenated to the foundational representation obtained in the first stage to adjust for the unmeasured confounding. To avoid overfitting caused by the small-scale RCT data in the second stage, we further propose a partial parameter initialization approach, rather than training a separate network. The superiority of our approach is validated on two public datasets with extensive experiments. The code is available at https://github.com/zhouchuanCN/KDD25-TSPF.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- Unveiling Extraneous Sampling Bias with Data Missing-Not-At-RandomChunyuan Zheng, Haocheng Yang, Haoxuan Li, Mengyue YangNeurIPS 2025 · 被引用 15 次
- Counterfactual Implicit Feedback ModelingChuan Zhou, Lina Yao, Haoxuan Li, Mingming GongNeurIPS 2025 · 被引用 8 次
- Addressing Correlated Latent Exogenous Variables in Debiased Recommender SystemsShuqiang Zhang, Yuchao Zhang, Jinkun Chen, Haochen SuiKDD 2025 · 被引用 4 次
- Proximity Matters: Local Proximity Enhanced Balancing for Treatment Effect EstimationHao Wang, Zhichao Chen, Zhaoran Liu, Xu Chen 等KDD 2025 · 被引用 4 次
- Unbiased Reward Modeling from Implicit Feedback for LLM AlignmentHao Wang, Haocheng Yang, Licheng Pan, Zhichao Chen 等ICML 2026 · 被引用 2 次
它引用的顶会 Paper16
- CLUB: A Contrastive Log-ratio Upper Bound of Mutual InformationPengyu Cheng, Weituo Hao, Shuyang Dai, Jiachang Liu 等ICML 2020 · 被引用 512 次
- Learning Disentangled Representations for CounterFactual RegressionNegar Hassanpour, Russell GreinerICLR 2020 · 被引用 176 次
- Estimating the Effects of Continuous-valued Interventions using Generative Adversarial NetworksIoana Bica, James Jordon, Mihaela van der SchaarNeurIPS 2020 · 被引用 137 次
- Treatment Effect Estimation with Disentangled Latent FactorsWeijia Zhang, Lin Liu, Jiuyong LiAAAI 2021 · 被引用 115 次
- Optimal Transport for Treatment Effect EstimationHao Wang, Jiajun Fan, Zhichao Chen, Haoxuan Li 等NeurIPS 2023 · 被引用 71 次
相关 Paper
- Quantifying Ignorance in Individual-Level Causal-Effect Estimates under Hidden ConfoundingAndrew Jesson, Sören Mindermann, Yarin Gal, Uri ShalitICML 2021 · 被引用 66 次
- A Non-parametric Direct Learning Approach to Heterogeneous Treatment Effect Estimation under Unmeasured ConfoundingXinhai Zhang, Xingye QiaoNeurIPS 2024 · 被引用 1 次
- Instrumental Variable Regression with Confounder BalancingAnpeng Wu, Kun Kuang, Bo Li, Fei WuICML 2022 · 被引用 31 次
- Conditional Instrumental Variable Regression with Representation Learning for Causal InferenceDebo Cheng, Ziqi Xu, Jiuyong Li, Lin Liu 等ICLR 2024 · 被引用 14 次
- A Neural Mean Embedding Approach for Back-door and Front-door AdjustmentLiyuan Xu, Arthur GrettonICLR 2023
