Self-Paced Robust Learning for Leveraging Clean Labels in Noisy Data
Xuchao Zhang, Xian Wu, Fanglan Chen, Liang Zhao, Chang-Tien Lu
Abstract
The success of training accurate models strongly depends on the availability of a sufficient collection of precisely labeled data. However, real-world datasets contain erroneously labeled data samples that substantially hinder the performance of machine learning models. Meanwhile, well-labeled data is usually expensive to obtain and only a limited amount is available for training. In this paper, we consider the problem of training a robust model by using large-scale noisy data in conjunction with a small set of clean data. To leverage the information contained via the clean labels, we propose a novel self-paced robust learning algorithm (SPRL) that trains the model in a process from more reliable (clean) data instances to less reliable (noisy) ones under the supervision of well-labeled data. The self-paced learning process hedges the risk of selecting corrupted data into the training set. Moreover, theoretical analyses on the convergence of the proposed algorithm are provided under mild assumptions. Extensive experiments on synthetic and real-world datasets demonstrate that our proposed approach can achieve a considerable improvement in effectiveness and robustness to existing methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers4
- Learning with Instance-Dependent Label Noise: A Sample Sieve ApproachHao Cheng, Zhaowei Zhu, Xingyu Li, Yifei Gong et al.ICLR 2021 · 27 citations
- Balanced Self-Paced Learning for AUC MaximizationBin Gu, Chenkang Zhang, Huan Xiong, Heng HuangAAAI 2022 · 5 citations
- Denoising Multi-Similarity Formulation: A Self-Paced Curriculum-Driven Approach for Robust Metric LearningChenkang Zhang, Lei Luo, Bin GuAAAI 2023 · 4 citations
- Incompatibility Clustering as a Defense Against Backdoor Poisoning AttacksCharles Jin, Melinda Sun, Martin C. RinardICLR 2023 · 1 citation
Related papers
- Sample-wise Label Confidence Incorporation for Learning with Noisy LabelsChanho Ahn, Kikyung Kim, Ji-Won Baek, Jongin Lim et al.ICCV 2023 · 11 citations
- A Model-Agnostic Approach for Learning with Noisy Labels of Arbitrary DistributionsShuang Hao, Peng Li, Renzhi Wu, Xu ChuICDE 2022 · 2 citations
- Robust Data Pruning under Label Noise via Maximizing Re-labeling AccuracyDongmin Park, Seola Choi, Doyoung Kim, Hwanjun Song et al.NeurIPS 2023 · 42 citations
- Scalable Penalized Regression for Noise Detection in Learning with Noisy LabelsYikai Wang, Xinwei Sun, Yanwei FuCVPR 2022 · 37 citations
- Error-Bounded Correction of Noisy LabelsSongzhu Zheng, Pengxiang Wu, Aman Goswami, Mayank Goswami et al.ICML 2020 · 153 citations
