Maximum Likelihood Estimation is All You Need for Well-Specified Covariate Shift
Jiawei Ge, Shange Tang, Jianqing Fan, Cong Ma, Chi Jin
摘要
A key challenge of modern machine learning systems is to achieve Out-of-Distribution (OOD) generalization -- generalizing to target data whose distribution differs from that of source data. Despite its significant importance, the fundamental question of ``what are the most effective algorithms for OOD generalization'' remains open even under the standard setting of covariate shift. This paper addresses this fundamental question by proving that, surprisingly, classical Maximum Likelihood Estimation (MLE) purely using source data (without any modification) achieves the minimax optimality for covariate shift under the well-specified setting. That is, no algorithm performs better than MLE in this setting (up to a constant factor), justifying MLE is all you need. Our result holds for a very rich class of parametric models, and does not require any boundedness condition on the density ratio. We illustrate the wide applicability of our framework by instantiating it to three concrete examples -- linear regression, logistic regression, and phase retrieval. This paper further complement the study by proving that, under the misspecified setting, MLE is no longer the optimal choice, whereas Maximum Weighted Likelihood Estimator (MWLE) emerges as minimax optimal in certain scenarios.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- High-Dimensional Kernel Methods under Covariate Shift: Data-Dependent Implicit RegularizationYihang Chen, Fanghui Liu, Taiji Suzuki, Volkan CevherICML 2024 · 被引用 5 次
- When Shift Happens - Confounding Is to BlameAbbavaram Gowtham Reddy, Celia Rubio-Madrigal, Rebekka Burkholz, Krikamol MuandetICLR 2026 · 被引用 5 次
- Mixed-Sample SGD: an End-to-end Analysis of Supervised Transfer LearningYuyang Deng, Samory KpotufeNeurIPS 2025 · 被引用 1 次
- Benign Overfitting in Out-of-Distribution Generalization of Linear ModelsShange Tang, Jiayun Wu, Jianqing Fan, Chi JinICLR 2025
它引用的顶会 Paper4
- The Risks of Invariant Risk MinimizationElan Rosenfeld, Pradeep Kumar Ravikumar, Andrej RisteskiICLR 2021 · 被引用 356 次
- Near-Optimal Linear Regression under Distribution ShiftQi Lei, Wei Hu, Jason D. LeeICML 2021 · 被引用 45 次
- A new similarity measure for covariate shift with applications to nonparametric regressionReese Pathak, Cong Ma, Martin J. WainwrightICML 2022 · 被引用 40 次
- Minimax Lower Bounds for Transfer Learning with Linear and One-hidden Layer Neural NetworksSeyed Mohammadreza Mousavi Kalan, Zalan Fabian, Salman Avestimehr, Mahdi SoltanolkotabiNeurIPS 2020 · 被引用 37 次
相关 Paper
- Towards a Unified Analysis of Kernel-based Methods Under Covariate ShiftXingdong Feng, Xin He, Caixing Wang, Chao Wang 等NeurIPS 2023 · 被引用 17 次
- A Theoretical Analysis on Independence-driven Importance Weighting for Covariate-shift GeneralizationRenzhe Xu, Xingxuan Zhang, Zheyan Shen, Tong Zhang 等ICML 2022 · 被引用 36 次
- Open Set Label Shift with Test Time Out-of-Distribution ReferenceChangkun Ye, Russell Tsuchida, Lars Petersson, Nick BarnesCVPR 2025
- Feed Two Birds with One Scone: Exploiting Wild Data for Both Out-of-Distribution Generalization and DetectionHaoyue Bai, Gregory Canal, Xuefeng Du, Jeongyeol Kwon 等ICML 2023 · 被引用 67 次
- Double-Weighting for Covariate Shift AdaptationJosé Ignacio Segovia-Martín, Santiago Mazuelas, Anqi LiuICML 2023 · 被引用 9 次
