A Unified View of Label Shift Estimation
Saurabh Garg, Yifan Wu, Sivaraman Balakrishnan, Zachary C. Lipton
摘要
Label shift describes the setting where although the label distribution might change between the source and target domains, the class-conditional probabilities (of data given a label) do not. There are two dominant approaches for estimating the label marginal. BBSE, a moment-matching approach based on confusion matrices, is provably consistent and provides interpretable error bounds. However, a maximum likelihood estimation approach, which we call MLLS, dominates empirically. In this paper, we present a unified view of the two methods and the first theoretical characterization of the likelihood-based estimator. Our contributions include (i) conditions for consistency of MLLS, which include calibration of the classifier and a confusion matrix invertibility condition that BBSE also requires; (ii) a unified view of the methods, casting the confusion matrix as roughly equivalent to MLLS for a particular choice of calibration method; and (iii) a decomposition of MLLS's finite-sample error into terms reflecting the impacts of miscalibration and estimation error. Our analysis attributes BBSE's statistical inefficiency to a loss of information due to coarse calibration. We support our findings with experiments on both synthetic data and the MNIST and CIFAR10 image recognition datasets.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper56
- WILDS: A Benchmark of in-the-Wild Distribution ShiftsPang Wei Koh, Shiori Sagawa, Henrik Marklund, Sang Michael Xie 等ICML 2021 · 被引用 1,773 次
- Self-Attention Between Datapoints: Going Beyond Individual Input-Output Pairs in Deep LearningJannik Kossen, Neil Band, Clare Lyle, Aidan N. Gomez 等NeurIPS 2021 · 被引用 180 次
- Leveraging unlabeled data to predict out-of-distribution performanceSaurabh Garg, Sivaraman Balakrishnan, Zachary Chase Lipton, Behnam Neyshabur 等ICLR 2022 · 被引用 160 次
- Maximum Likelihood with Bias-Corrected Calibration is Hard-To-Beat at Label Shift AdaptationAmr Alexandari, Anshul Kundaje, Avanti ShrikumarICML 2020 · 被引用 123 次
- Distribution-free binary classification: prediction sets, confidence intervals and calibrationChirag Gupta, Aleksandr Podkopaev, Aaditya RamdasNeurIPS 2020 · 被引用 105 次
相关 Paper
- LaSCal: Label-Shift Calibration without target labelsTeodora Popordanoska, Gorjan Radevski, Tinne Tuytelaars, Matthew B. BlaschkoNeurIPS 2024 · 被引用 12 次
- Expectation Consistency Loss: Rethink Confidence Calibration under Covariate ShiftJinzong Dong, Zhaohui Jiang, Bo YangICML 2026
- ELSA: Efficient Label Shift Adaptation through the Lens of Semiparametric ModelsQinglong Tian, Xin Zhang, Jiwei ZhaoICML 2023 · 被引用 12 次
- Label Shift Correction via Bidirectional Marginal Distribution MatchingRuidong Fan, Xiao Ouyang, Hong Tao, Chenping HouKDD 2024 · 被引用 1 次
- Graph-Smoothed Bayesian Black-Box Shift Estimator and Its Information GeometryMasanari KimuraNeurIPS 2025 · 被引用 1 次
