Imbalance-Aware Uplift Modeling for Observational Data
Xuanying Chen, Zhining Liu, Li Yu, Liuyi Yao, Wenpeng Zhang, Yi Dong, Lihong Gu, Xiaodong Zeng, Yize Tan, Jinjie Gu
摘要
Uplift modeling aims to model the incremental impact of a treatment on an individual outcome, which has attracted great interests of researchers and practitioners from different communities. Existing uplift modeling methods rely on either the data collected from randomized controlled trials (RCTs) or the observational data which is more realistic. However, we notice that on the observational data, it is often the case that only a small number of subjects receive treatment, but finally infer the uplift on a much large group of subjects. Such highly imbalanced data is common in various fields such as marketing and medical treatment but it is rarely handled by existing works. In this paper, we theoretically and quantitatively prove that the existing representative methods, transformed outcome (TOM) and doubly robust (DR), suffer from large bias and deviation on highly imbalanced datasets with skewed propensity scores, mainly because they are proportional to the reciprocal of the propensity score. To reduce the bias and deviation of uplift modeling with an imbalanced dataset, we propose an imbalance-aware uplift modeling (IAUM) method via constructing a robust proxy outcome, which adaptively combines the doubly robust estimator and the imputed treatment effects based on the propensity score. We theoretically prove that IAUM can obtain a better bias-variance trade-off than existing methods on a highly imbalanced dataset. We conduct extensive experiments on a synthetic dataset and two real-world datasets, and the experimental results well demonstrate the superiority of our method over state-of-the-art.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
相关 Paper
- Invariant Deep Uplift Modeling for Incentive Assignment in Online Marketing via Probability of Necessity and SufficiencyZexu Sun, Qiyu Han, Hao Yang, Anpeng Wu 等ICML 2025
- Multiple Robust Learning for RecommendationHaoxuan Li, Quanyu Dai, Yuru Li, Yan Lyu 等AAAI 2023 · 被引用 48 次
- Unified Minimax Optimization Framework for Propensity Score Estimation in Debiased RecommendationChunyuan Zheng, Haocheng Yang, Jinkun Chen, Shufeng Zhang 等AAAI 2026 · 被引用 2 次
- Continuous Treatment Effects with Surrogate OutcomesZhenghao Zeng, David Arbour, Avi Feller, Raghavendra Addanki 等ICML 2024 · 被引用 4 次
- Graph Neural Network with Two Uplift Estimators for Label-Scarcity Individual Uplift ModelingDingyuan Zhu, Daixin Wang, Zhiqiang Zhang, Kun Kuang 等WWW 2023 · 被引用 3 次
