Generalized Jensen-Shannon Divergence Loss for Learning with Noisy Labels
Erik Englesson, Hossein Azizpour
Abstract
Prior works have found it beneficial to combine provably noise-robust loss functions e.g., mean absolute error (MAE) with standard categorical loss function e.g. cross entropy (CE) to improve their learnability. Here, we propose to use Jensen-Shannon divergence as a noise-robust loss function and show that it interestingly interpolate between CE and MAE with a controllable mixing parameter. Furthermore, we make a crucial observation that CE exhibit lower consistency around noisy data points. Based on this observation, we adopt a generalized version of the Jensen-Shannon divergence for multiple distributions to encourage consistency around data points. Using this loss function, we show state-of-the-art results on both synthetic (CIFAR), and real-world (e.g., WebVision) noise with varying noise rates.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext bc1c019d-8226-42f6-95c5-1ae1aa78eccdCited by top-tier papers36
- To Smooth or Not? When Label Smoothing Meets Noisy LabelsJiaheng Wei, Hangyu Liu, Tongliang Liu, Gang Niu et al.ICML 2022 · 104 citations
- Learning with Neighbor Consistency for Noisy LabelsAhmet Iscen, Jack Valmadre, Anurag Arnab, Cordelia SchmidCVPR 2022 · 79 citations
- DNA: Domain Generalization with Diversified Neural AveragingXu Chu, Yujie Jin, Wenwu Zhu, Yasha Wang et al.ICML 2022 · 41 citations
- Bridging the Gap Between Promise and Performance for Microscaling FP4 QuantizationVage Egiazarian, Roberto L. Castro, Denis Kuznedelev, Andrei Panferov et al.ICLR 2026 · 38 citations
- FedDiv: Collaborative Noise Filtering for Federated Learning with Noisy LabelsJichang Li, Guanbin Li, Hui Cheng, Zicheng Liao et al.AAAI 2024 · 37 citations
Builds on11
- RandAugment: Practical Automated Data Augmentation with a Reduced Search SpaceEkin Dogus Cubuk, Barret Zoph, Jonathon Shlens, Quoc LeNeurIPS 2020 · 4,453 citations
- AugMix: A Simple Data Processing Method to Improve Robustness and UncertaintyDan Hendrycks, Norman Mu, Ekin Dogus Cubuk, Barret Zoph et al.ICLR 2020 · 1,572 citations
- DivideMix: Learning with Noisy Labels as Semi-supervised LearningJunnan Li, Richard Socher, Steven C. H. HoiICLR 2020 · 1,326 citations
- Symmetric Cross Entropy for Robust Learning With Noisy LabelsYisen Wang, Xingjun Ma, Zaiyi Chen, Yuan Luo et al.ICCV 2019 · 1,125 citations
- Early-Learning Regularization Prevents Memorization of Noisy LabelsSheng Liu, Jonathan Niles-Weed, Narges Razavian, Carlos Fernandez-GrandaNeurIPS 2020 · 798 citations
Related papers
- Label Distributionally Robust Losses for Multi-class Classification: Consistency, Robustness and AdaptivityDixian Zhu, Yiming Ying, Tianbao YangICML 2023 · 15 citations
- Normalized Loss Functions for Deep Learning with Noisy LabelsXingjun Ma, Hanxun Huang, Yisen Wang, Simone Romano et al.ICML 2020 · 547 citations
- When Optimizing f-Divergence is Robust with Label NoiseJiaheng Wei, Yang LiuICLR 2021 · 64 citations
- On Learning Contrastive Representations for Learning with Noisy LabelsLi Yi, Sheng Liu, Qi She, A. Ian McLeod et al.CVPR 2022 · 68 citations
- Learning from Noisy Labels with Complementary Loss FunctionsDeng-Bao Wang, Yong Wen, Lujia Pan, Min-Ling ZhangAAAI 2021 · 40 citations
