Learning from Noisy Labels with No Change to the Training Process
Mingyuan Zhang, Jane H. Lee, Shivani Agarwal
摘要
There has been much interest in recent years in developing learning algorithms that can learn accurate classifiers from data with noisy labels. A widely-studied noise model is that of classconditional noise (CCN), wherein a label y is flipped to a label y with some associated noise probability that depends on both y and y. In the multiclass setting, all previously proposed algorithms under the CCN model involve changing the training process, by introducing a 'noisecorrection' to the surrogate loss to be minimized over the noisy training examples. In this paper, we show that this is really unnecessary: one can simply perform class probability estimation (CPE) on the noisy examples, e.g. using a standard (multiclass) logistic regression algorithm, and then apply noise-correction only in the final prediction step. This means that the training algorithm itself does not need any change, and one can simply use standard off-the-shelf implementations with no modification to the code for training. Our approach can handle general multiclass loss matrices, including the usual 0-1 loss but also other losses such as those used for ordinal regression problems. We also provide a quantitative regret transfer bound, which bounds the target regret on the true distribution in terms of the CPE regret on the noisy distribution; in doing so, we extend the notion of strong properness introduced for binary losses by Agarwal (2014) to the multiclass case. Our bound suggests that the sample complexity of learning under CCN increases as the noise matrix approaches singularity. We also provide fixes and potential improvements for noise estimation methods that involve computing anchor points. Our experiments confirm our theoretical findings.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper14
- Detecting Corrupted Labels Without Training a Model to PredictZhaowei Zhu, Zihao Dong, Yang LiuICML 2022 · 被引用 84 次
- Beyond Images: Label Noise Transition Matrix Estimation for Tasks with Lower-Quality FeaturesZhaowei Zhu, Jialu Wang, Yang LiuICML 2022 · 被引用 43 次
- Learning from Label Proportions by Learning with Label NoiseJianxin Zhang, Yutong Wang, Clayton ScottNeurIPS 2022 · 被引用 41 次
- Delving into Noisy Label Detection with Clean DataChenglin Yu, Xinsong Ma, Weiwei LiuICML 2023 · 被引用 19 次
- On Learning Latent Models with Multi-Instance Weak SupervisionKaifu Wang, Efthymia Tsamoura, Dan RothNeurIPS 2023 · 被引用 19 次
它引用的顶会 Paper3
- Dual T: Reducing Estimation Error for Transition Matrix in Label-noise LearningYu Yao, Tongliang Liu, Bo Han, Mingming Gong 等NeurIPS 2020 · 被引用 297 次
- Learning with Bounded Instance and Label-dependent Label NoiseJiacheng Cheng, Tongliang Liu, Kotagiri Ramamohanarao, Dacheng TaoICML 2020 · 被引用 162 次
- Intra Order-preserving Functions for Calibration of Multi-Class Neural NetworksAmir Rahimi, Amirreza Shaban, Ching-An Cheng, Richard Hartley 等NeurIPS 2020 · 被引用 96 次
相关 Paper
- Estimating Noise Transition Matrix with Label Correlations for Noisy Multi-Label LearningShikun Li, Xiaobo Xia, Hansong Zhang, Yibing Zhan 等NeurIPS 2022 · 被引用 95 次
- From Noisy Prediction to True Label: Noisy Prediction Calibration via Generative ModelHeeSun Bae, Seungjae Shin, Byeonghu Na, JoonHo Jang 等ICML 2022 · 被引用 29 次
- Learning Noise Transition Matrix from Only Noisy Labels via Total Variation RegularizationYivan Zhang, Gang Niu, Masashi SugiyamaICML 2021 · 被引用 107 次
- Resurfacing the Instance-only Dependent Label Noise Model through Loss CorrectionMustafa Enes Aydın, Maarten De Vos, Alexander BertrandICLR 2026
- Estimating Instance-dependent Bayes-label Transition Matrix using a Deep Neural NetworkShuo Yang, Erkun Yang, Bo Han, Yang Liu 等ICML 2022 · 被引用 59 次
