Regroup Median Loss for Combating Label Noise
Fengpeng Li, Kemou Li, Jinyu Tian, Jiantao Zhou
Abstract
The deep model training procedure requires large-scale datasets of annotated data. Due to the difficulty of annotating a large number of samples, label noise caused by incorrect annotations is inevitable, resulting in low model performance and poor model generalization. To combat label noise, current methods usually select clean samples based on the small-loss criterion and use these samples for training. Due to some noisy samples similar to clean ones, these small-loss criterion-based methods are still affected by label noise. To address this issue, in this work, we propose Regroup Median Loss (RML) to reduce the probability of selecting noisy samples and correct losses of noisy samples. RML randomly selects samples with the same label as the training samples based on a new loss processing method. Then, we combine the stable mean loss and the robust median loss through a proposed regrouping strategy to obtain robust loss estimation for noisy samples. To further improve the model performance against label noise, we propose a new sample selection strategy and build a semi-supervised method based on RML. Compared to state-of-the-art methods, for both the traditionally trained and semi-supervised models, RML achieves a significant improvement on synthetic and complex real-world datasets. The source is at https://github.com/Feng-peng-Li/Regroup-Loss-Median-to-Combat-Label-Noise.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers8
- From Pretrain to Pain: Adversarial Vulnerability of Video Foundation Models Without Task KnowledgeHui Lu, Yi Yu, Song Xia, Yiming Yang et al.AAAI 2026 · 8 citations
- Enhancing Sample Selection Against Label Noise by Cutting Mislabeled Easy ExamplesSuqin Yuan, Lei Feng, Bo Han, Tongliang LiuNeurIPS 2025 · 5 citations
- Learning with Open-world Noisy Data via Class-independent Margin in Dual Representation SpaceLinchao Pan, Can Gao, Jie Zhou, Jinbao WangAAAI 2025 · 1 citation
- Editprint: General Digital Image Forensics via Editing Fingerprint with Self-Augmentation TrainingHaiwei Wu, Kemou Li, Yuanman Li, Jiantao ZhouCVPR 2026
- DREAM: Dual-Standard Semantic Homogeneity with Dynamic Optimization for Graph Learning with Label NoiseYusheng Zhao, Jiaye Xie, Qixin Zhang, Weizhi Zhang et al.ICML 2026
Builds on14
- DivideMix: Learning with Noisy Labels as Semi-supervised LearningJunnan Li, Richard Socher, Steven C. H. HoiICLR 2020 · 1,326 citations
- Early-Learning Regularization Prevents Memorization of Noisy LabelsSheng Liu, Jonathan Niles-Weed, Narges Razavian, Carlos Fernandez-GrandaNeurIPS 2020 · 798 citations
- Part-dependent Label Noise: Towards Instance-dependent Label NoiseXiaobo Xia, Tongliang Liu, Bo Han, Nannan Wang et al.NeurIPS 2020 · 329 citations
- Understanding and Improving Early Stopping for Learning with Noisy LabelsYingbin Bai, Erkun Yang, Bo Han, Yanhua Yang et al.NeurIPS 2021 · 307 citations
- Selective-Supervised Contrastive Learning with Noisy LabelsShikun Li, Xiaobo Xia, Shiming Ge, Tongliang LiuCVPR 2022 · 201 citations
Related papers
- SELF: Learning to Filter Noisy Labels with Self-EnsemblingDuc Tam Nguyen, Chaithanya Kumar Mummadi, Thi-Phuong-Nhung Ngo, Thi Hoai Phuong Nguyen et al.ICLR 2020 · 354 citations
- Sample Selection with Uncertainty of Losses for Learning with Noisy LabelsXiaobo Xia, Tongliang Liu, Bo Han, Mingming Gong et al.ICLR 2022 · 139 citations
- USDNL: Uncertainty-Based Single Dropout in Noisy Label LearningYuanzhuo Xu, Xiaoguang Niu, Jie Yang, Steve Drew et al.AAAI 2023 · 9 citations
- Jo-SRC: A Contrastive Approach for Combating Noisy LabelsYazhou Yao, Zeren Sun, Chuanyi Zhang, Fumin Shen et al.CVPR 2021
- Which is Better for Learning with Noisy Labels: The Semi-supervised Method or Modeling Label Noise?Yu Yao, Mingming Gong, Yuxuan Du, Jun Yu et al.ICML 2023 · 15 citations
