Robust Curriculum Learning: from clean label detection to noisy label self-correction
Tianyi Zhou, Shengjie Wang, Jeff A. Bilmes
摘要
Neural network training can easily overfit noisy labels resulting in poor generalization performance. Existing methods address this problem by (1) filtering out the noisy data and only using the clean data for training or (2) relabeling the noisy data by the model during training or by another model trained only on a clean dataset. However, the former does not leverage the features' information of wrongly-labeled data, while the latter may produce wrong pseudo-labels for some data and introduce extra noises. In this paper, we propose a smooth transition and interplay between these two strategies as a curriculum that selects training samples dynamically. In particular, we start with learning from clean data and then gradually move to learn noisy-labeled data with pseudo labels produced by a time-ensemble of the model and data augmentations. Instead of using the instantaneous loss computed at the current step, our data selection is based on the dynamics of both the loss and output consistency for each sample across historical steps and different data augmentations, resulting in more precise detection of both clean labels and correct pseudo labels. On multiple benchmarks of noisy labels, we show that our curriculum learning strategy can significantly improve the test accuracy without any auxiliary model or extra clean data.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper36
- Adaptive Early-Learning Correction for Segmentation from Noisy AnnotationsSheng Liu, Kangning Liu, Weicheng Zhu, Yiqiu Shen 等CVPR 2022 · 被引用 109 次
- Combating Noisy Labels with Sample Selection by Mining High-Discrepancy ExamplesXiaobo Xia, Bo Han, Yibing Zhan, Jun Yu 等ICCV 2023 · 被引用 72 次
- Sketching without Worrying: Noise-Tolerant Sketch-Based Image RetrievalAyan Kumar Bhunia, Subhadeep Koley, Abdullah Faiz Ur Rahman Khilji, Aneeshan Sain 等CVPR 2022 · 被引用 53 次
- Robust Data Pruning under Label Noise via Maximizing Re-labeling AccuracyDongmin Park, Seola Choi, Doyoung Kim, Hwanjun Song 等NeurIPS 2023 · 被引用 42 次
- Early Stopping Against Label Noise Without Validation DataSuqin Yuan, Lei Feng, Tongliang LiuICLR 2024 · 被引用 39 次
相关 Paper
- SuperLoss: A Generic Loss for Robust Curriculum LearningThibault Castells, Philippe Weinzaepfel, Jérôme RevaudNeurIPS 2020 · 被引用 96 次
- SELF: Learning to Filter Noisy Labels with Self-EnsemblingDuc Tam Nguyen, Chaithanya Kumar Mummadi, Thi-Phuong-Nhung Ngo, Thi Hoai Phuong Nguyen 等ICLR 2020 · 被引用 354 次
- Time-Consistent Self-Supervision for Semi-Supervised LearningTianyi Zhou, Shengjie Wang, Jeff A. BilmesICML 2020 · 被引用 58 次
- TrainRef: Curating Data with Label Distribution and Minimal Reference for Accurate Prediction and Reliable ConfidenceMurong Ma, Ruofan Liu, Yun Lin, Zhiyong Huang 等ICLR 2026
- C-SFDA: A Curriculum Learning Aided Self-Training Framework for Efficient Source Free Domain AdaptationNazmul Karim, Niluthpol Chowdhury Mithun, Abhinav Rajvanshi, Han-Pang Chiu 等CVPR 2023
