Unveiling Extraneous Sampling Bias with Data Missing-Not-At-Random
Chunyuan Zheng, Haocheng Yang, Haoxuan Li, Mengyue Yang
摘要
Selection bias poses a widely recognized challenge for unbiased evaluation and learning in many industrial scenarios. For example, in recommender systems, it arises from the users' selective interactions with items. Recently, doubly robust and its variants have been widely studied to achieve debiased learning of prediction models, however, all of them consider a simple exact matching scenario, i.e., the units (such as user-item pairs in a recommender system) are the same between the training and test sets. In practice, there may be limited or even no overlap in units between the training and test. In this paper, we consider a more practical scenario: the joint distribution of the feature and rating is the same in the training and test sets. Theoretical analysis shows that the previous DR estimator is biased even if the imputed errors and learned propensities are correct in this scenario. In addition, we propose a novel super-population doubly robust estimator (SuperDR), which can achieve a more accurate estimation and desirable generalization error bound compared to the existing DR estimators, and extend the joint learning algorithm for training the prediction and imputation models. We conduct extensive experiments on three real-world datasets, including a large-scale industrial dataset, to show the effectiveness of our method. The code is available at https://github.com/ChunyuanZheng/neurips-25-SuperDR.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Counterfactual Implicit Feedback ModelingChuan Zhou, Lina Yao, Haoxuan Li, Mingming GongNeurIPS 2025 · 被引用 8 次
- Unbiased Reward Modeling from Implicit Feedback for LLM AlignmentHao Wang, Haocheng Yang, Licheng Pan, Zhichao Chen 等ICML 2026 · 被引用 2 次
- Uplift Modeling with Delayed Feedback: Identifiability and AlgorithmsChunyuan Zheng, Anpeng Wu, Chuan Zhou, Taojun Hu 等AAAI 2026 · 被引用 2 次
- Optimizing Marketing Subsidies via Counterfactual Learning with Asymmetric Reward FunctionXiang Li, Yanghao Xiao, Chunyuan Zheng, Qian Zou 等SIGIR 2026 · 被引用 1 次
- Debiased Recommendation Beyond the Positive Propensity AssumptionYanghao Xiao, Hao Wang, Xiang Li, Qian Zou 等SIGIR 2026
它引用的顶会 Paper31
- Causal Intervention for Leveraging Popularity Bias in RecommendationYang Zhang, Fuli Feng, Xiangnan He, Tianxin Wei 等SIGIR 2021 · 被引用 431 次
- A General Knowledge Distillation Framework for Counterfactual Recommendation via Uniform DataDugang Liu, Pengxiang Cheng, Zhenhua Dong, Xiuqiang He 等SIGIR 2020 · 被引用 188 次
- AutoDebias: Learning to Debias for RecommendationJiawei Chen, Hande Dong, Yang Qiu, Xiangnan He 等SIGIR 2021 · 被引用 167 次
- Information Theoretic Counterfactual Learning from Missing-Not-At-Random FeedbackZifeng Wang, Xi Chen, Rui Wen, Shao-Lun Huang 等NeurIPS 2020 · 被引用 95 次
- Asymmetric Tri-training for Debiasing Missing-Not-At-Random Explicit FeedbackYuta SaitoSIGIR 2020 · 被引用 90 次
相关 Paper
- Doubly Calibrated Estimator for Recommendation on Data Missing Not at RandomWonbin Kweon, Hwanjo YuWWW 2024 · 被引用 23 次
- Multiple Robust Learning for RecommendationHaoxuan Li, Quanyu Dai, Yuru Li, Yan Lyu 等AAAI 2023 · 被引用 48 次
- Relaxing the Accurate Imputation Assumption in Doubly Robust Learning for Debiased Collaborative FilteringHaoxuan Li, Chunyuan Zheng, Shuyi Wang, Kunhan Wu 等ICML 2024 · 被引用 25 次
- StableDR: Stabilized Doubly Robust Learning for Recommendation on Data Missing Not at RandomHaoxuan Li, Chunyuan Zheng, Peng WuICLR 2023 · 被引用 12 次
- TDR-CL: Targeted Doubly Robust Collaborative Learning for Debiased RecommendationsHaoxuan Li, Yan Lyu, Chunyuan Zheng, Peng WuICLR 2023 · 被引用 14 次
