Uncertainty-Aware Learning against Label Noise on Imbalanced Datasets
Yingsong Huang, Bing Bai, Shengwei Zhao, Kun Bai, Fei Wang
Abstract
Learning against label noise is a vital topic to guarantee a reliable performance for deep neural networks. Recent research usually refers to dynamic noise modeling with model output probabilities and loss values, and then separates clean and noisy samples. These methods have gained notable success. However, unlike cherry-picked data, existing approaches often cannot perform well when facing imbalanced datasets, a common scenario in the real world. We thoroughly investigate this phenomenon and point out two major issues that hinder the performance, i.e., inter-class loss distribution discrepancy and misleading predictions due to uncertainty. The first issue is that existing methods often perform class-agnostic noise modeling. However, loss distributions show a significant discrepancy among classes under class imbalance, and class-agnostic noise modeling can easily get confused with noisy samples and samples in minority classes. The second issue refers to that models may output misleading predictions due to epistemic uncertainty and aleatoric uncertainty, thus existing methods that rely solely on the output probabilities may fail to distinguish confident samples. Inspired by our observations, we propose an Uncertainty-aware Label Correction framework (ULC) to handle label noise on imbalanced datasets. First, we perform epistemic uncertainty-aware classspecific noise modeling to identify trustworthy clean samples and refine/discard highly confident true/corrupted labels. Then, we introduce aleatoric uncertainty in the subsequent learning process to prevent noise accumulation in the label noise modeling process. We conduct experiments on several synthetic and real-world datasets. The results demonstrate the effectiveness of the proposed method, especially on imbalanced datasets.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext eda4dae8-c9cb-490c-979a-86ade341ca5cCited by top-tier papers12
- Label-Noise Learning with Intrinsically Long-Tailed DataYang Lu, Yiliang Zhang, Bo Han, Yiu-Ming Cheung et al.ICCV 2023 · 32 citations
- Dirichlet-based Per-Sample Weighting by Transition Matrix for Noisy Label LearningHeeSun Bae, Seungjae Shin, Byeonghu Na, Il-Chul MoonICLR 2024 · 10 citations
- USDNL: Uncertainty-Based Single Dropout in Noisy Label LearningYuanzhuo Xu, Xiaoguang Niu, Jie Yang, Steve Drew et al.AAAI 2023 · 9 citations
- Diffusion Epistemic Uncertainty with Asymmetric Learning for Diffusion-Generated Image DetectionYingsong Huang, Hui Guo, Jing Huang, Bing Bai et al.ICCV 2025 · 4 citations
- Noisy Multi-Label Learning through Co-Occurrence-Aware DiffusionSenyu Hou, Yuru Ren, Gaoxia Jiang, Wenjian WangNeurIPS 2025 · 3 citations
Builds on8
- FixMatch: Simplifying Semi-Supervised Learning with Consistency and ConfidenceKihyuk Sohn, David Berthelot, Nicholas Carlini, Zizhao Zhang et al.NeurIPS 2020 · 5,129 citations
- Decoupling Representation and Classifier for Long-Tailed RecognitionBingyi Kang, Saining Xie, Marcus Rohrbach, Zhicheng Yan et al.ICLR 2020 · 1,496 citations
- DivideMix: Learning with Noisy Labels as Semi-supervised LearningJunnan Li, Richard Socher, Steven C. H. HoiICLR 2020 · 1,326 citations
- Early-Learning Regularization Prevents Memorization of Noisy LabelsSheng Liu, Jonathan Niles-Weed, Narges Razavian, Carlos Fernandez-GrandaNeurIPS 2020 · 798 citations
- In Defense of Pseudo-Labeling: An Uncertainty-Aware Pseudo-label Selection Framework for Semi-Supervised LearningMamshad Nayeem Rizve, Kevin Duarte, Yogesh S. Rawat, Mubarak ShahICLR 2021 · 630 citations
Related papers
- PNP: Robust Learning from Noisy Labels by Probabilistic Noise PredictionZeren Sun, Fumin Shen, Dan Huang, Qiong Wang et al.CVPR 2022 · 79 citations
- Noise-Robust Learning from Multiple Unsupervised Sources of Inferred LabelsAmila Silva, Ling Luo, Shanika Karunasekera, Christopher LeckieAAAI 2022 · 11 citations
- Training Noise-Robust Deep Neural Networks via Meta-LearningZhen Wang, Guosheng Hu, Qinghua HuCVPR 2020
- Sample Selection with Uncertainty of Losses for Learning with Noisy LabelsXiaobo Xia, Tongliang Liu, Bo Han, Mingming Gong et al.ICLR 2022 · 139 citations
- Confidence-based Reliable Learning under Dual NoisesPeng Cui, Yang Yue, Zhijie Deng, Jun ZhuNeurIPS 2022 · 13 citations
