General Quantification of Covariate and Concept Shifts
Hongbo Chen, Li Xia
摘要
Generalization under distribution shift remains a core challenge in modern machine learning, yet existing learning bound theory is limited to narrow, idealized settings and is non-estimable from samples. In this paper, we bridge the gap between theory and practical applications. We first show that existing bounds become loose and nonestimable because their concept shift definition breaks when the source and target supports mismatch. Leveraging entropic optimal transport, we propose new support-agnostic definitions for covariate and concept shifts, and derive a novel unified error bound that applies to broad loss functions, label spaces, and stochastic labeling. We further develop estimators for these shifts with concentration guarantees, and the DataShifts algorithm, which can quantify distribution shifts and estimate the error bound in most applications -a rigorous and general tool for analyzing learning error under distribution shift.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper3
- Out-of-Distribution Generalization via Risk Extrapolation (REx)David Krueger, Ethan Caballero, Jörn-Henrik Jacobsen, Amy Zhang 等ICML 2021 · 被引用 1,163 次
- Towards a Theoretical Framework of Out-of-Distribution GeneralizationHaotian Ye, Chuanlong Xie, Tianle Cai, Ruichen Li 等NeurIPS 2021 · 被引用 159 次
- NICO++: Towards Better Benchmarking for Domain GeneralizationXingxuan Zhang, Yue He, Renzhe Xu, Han Yu 等CVPR 2023
相关 Paper
- Non-exchangeable Conformal Prediction with Optimal Transport: Tackling Distribution Shift with Unlabeled DataAlvaro H. C. Correia, Christos LouizosNeurIPS 2025 · 被引用 5 次
- Margin-aware Adversarial Domain Adaptation with Optimal TransportSofien Dhouib, Ievgen Redko, Carole LartizienICML 2020 · 被引用 17 次
- Universal generalization guarantees for Wasserstein distributionally robust modelsTam Le, Jérôme MalickICLR 2025
- A new similarity measure for covariate shift with applications to nonparametric regressionReese Pathak, Cong Ma, Martin J. WainwrightICML 2022 · 被引用 40 次
- Incorporating Importance Weighting in Optimal Transport Based Domain AlignmentOkan Koç, Alexander Soen, Shanglin Li, Masashi SugiyamaICML 2026
