When does Privileged information Explain Away Label Noise?
Guillermo Ortiz-Jiménez, Mark Collier, Anant Nawalgaria, Alexander Nicholas D'Amour, Jesse Berent, Rodolphe Jenatton, Efi Kokiopoulou
Abstract
Leveraging privileged information (PI), or features available during training but not at test time, has recently been shown to be an effective method for addressing label noise. However, the reasons for its effectiveness are not well understood. In this study, we investigate the role played by different properties of the PI in explaining away label noise. Through experiments on multiple datasets with real PI (CIFAR-N/H) and a new large-scale benchmark ImageNet-PI, we find that PI is most helpful when it allows networks to easily distinguish clean from mislabeled data, while enabling a learning shortcut to memorize the mislabeled examples. Interestingly, when PI becomes too predictive of the target label, PI methods often perform worse than their no-PI baselines. Based on these findings, we propose several enhancements to the state-of-the-art PI methods and demonstrate the potential of PI as a means of tackling label noise. Finally, we show how we can easily combine the resulting PI approaches with existing no-PI techniques designed to deal with label noise.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 10fe36c3-e398-499d-bef9-3696eefee263Cited by top-tier papers2
- Pi-DUAL: Using privileged information to distinguish clean from noisy labelsKe Wang, Guillermo Ortiz-Jiménez, Rodolphe Jenatton, Mark Collier et al.ICML 2024 · 7 citations
- Conformal Prediction with Corrupted Labels: Uncertain Imputation and Robust Re-weightingShai Feldman, Stephen Bates, Yaniv RomanoICLR 2026 · 5 citations
Builds on16
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Early-Learning Regularization Prevents Memorization of Noisy LabelsSheng Liu, Jonathan Niles-Weed, Narges Razavian, Carlos Fernandez-GrandaNeurIPS 2020 · 798 citations
- Gradient Starvation: A Learning Proclivity in Neural NetworksMohammad Pezeshki, Sékou-Oumar Kaba, Yoshua Bengio, Aaron C. Courville et al.NeurIPS 2021 · 378 citations
- Human Uncertainty Makes Classification More RobustJoshua C. Peterson, Ruairidh M. Battleday, Thomas L. Griffiths, Olga RussakovskyICCV 2019 · 362 citations
- Learning with Noisy Labels Revisited: A Study Using Real-World Human AnnotationsJiaheng Wei, Zhaowei Zhu, Hao Cheng, Tongliang Liu et al.ICLR 2022 · 338 citations
Related papers
- Transfer and Marginalize: Explaining Away Label Noise with Privileged InformationMark Collier, Rodolphe Jenatton, Effrosyni Kokiopoulou, Jesse BerentICML 2022 · 19 citations
- Beyond Synthetic Noise: Deep Learning on Controlled Noisy LabelsLu Jiang, Di Huang, Mason Liu, Weilong YangICML 2020 · 241 citations
- Noise Attention Learning: Enhancing Noise Robustness by Gradient ScalingYangdi Lu, Yang Bo, Wenbo HeNeurIPS 2022 · 13 citations
- Re-Labeling ImageNet: From Single to Multi-Labels, From Global to Localized LabelsSangdoo Yun, Seong Joon Oh, Byeongho Heo, Dongyoon Han et al.CVPR 2021
- Distilling Effective Supervision From Severe Label NoiseZizhao Zhang, Han Zhang, Sercan Ömer Arik, Honglak Lee et al.CVPR 2020
