Correlated Input-Dependent Label Noise in Large-Scale Image Classification
Mark Collier, Basil Mustafa, Efi Kokiopoulou, Rodolphe Jenatton, Jesse Berent
摘要
Large scale image classification datasets often contain noisy labels. We take a principled probabilistic approach to modelling input-dependent, also known as heteroscedastic, label noise in these datasets. We place a multivariate Normal distributed latent variable on the final hidden layer of a neural network classifier. The covariance matrix of this latent variable, models the aleatoric uncertainty due to label noise. We demonstrate that the learned covariance structure captures known sources of label noise between semantically similar and co-occurring classes. Compared to standard neural network training and other baselines, we show significantly improved accuracy on Imagenet ILSVRC 2012 79.3% (+ 2.6%), Imagenet-21k 47.0% (+ 1.1%) and JFT 64.7% (+ 1.6%). We set a new state-of-the-art result on WebVision 1.0 with 76.6% top-1 accuracy. These datasets range from over 1M to over 300M training examples and from 1k classes to more than 21k classes. Our method is simple to use, and we provide an implementation that is a drop-in replacement for the final fully-connected layer in a deep classifier.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper22
- Scaling Vision Transformers to 22 Billion ParametersMostafa Dehghani, Josip Djolonga, Basil Mustafa, Piotr Padlewski 等ICML 2023 · 被引用 848 次
- Epistemic Neural NetworksIan Osband, Zheng Wen, Seyed Mohammad Asghari, Vikranth Dwaracherla 等NeurIPS 2023 · 被引用 142 次
- Uncertainty Estimation by Fisher Information-based Evidential Deep LearningDanruo Deng, Guangyong Chen, Yang Yu, Furui Liu 等ICML 2023 · 被引用 82 次
- Learning with Neighbor Consistency for Noisy LabelsAhmet Iscen, Jack Valmadre, Anurag Arnab, Cordelia SchmidCVPR 2022 · 被引用 79 次
- Open-Vocabulary Instance Segmentation via Robust Cross-Modal Pseudo-LabelingDat Huynh, Jason Kuen, Zhe Lin, Jiuxiang Gu 等CVPR 2022 · 被引用 78 次
它引用的顶会 Paper7
- Bayesian Deep Learning and a Probabilistic Perspective of GeneralizationAndrew Gordon Wilson, Pavel IzmailovNeurIPS 2020 · 被引用 845 次
- Beyond Synthetic Noise: Deep Learning on Controlled Noisy LabelsLu Jiang, Di Huang, Mason Liu, Weilong YangICML 2020 · 被引用 241 次
- Efficient and Scalable Bayesian Neural Nets with Rank-1 FactorsMichael Dusenberry, Ghassen Jerfel, Yeming Wen, Yi-An Ma 等ICML 2020 · 被引用 239 次
- Stochastic Segmentation Networks: Modelling Spatially Correlated Aleatoric UncertaintyMiguel Monteiro, Loïc Le Folgoc, Daniel Coelho de Castro, Nick Pawlowski 等NeurIPS 2020 · 被引用 153 次
- Combining Ensembles and Data Augmentation Can Harm Your CalibrationYeming Wen, Ghassen Jerfel, Rafael Muller, Michael W. Dusenberry 等ICLR 2021 · 被引用 72 次
相关 Paper
- Heteroskedastic and Imbalanced Deep Learning with Adaptive RegularizationKaidi Cao, Yining Chen, Junwei Lu, Nikos Aréchiga 等ICLR 2021 · 被引用 20 次
- Neural Dependencies Emerging from Learning Massive CategoriesRuili Feng, Kecheng Zheng, Kai Zhu, Yujun Shen 等CVPR 2023
- Enhancing Noise-Robust Losses for Large-Scale Noisy Data LearningMax Staats, Matthias Thamm, Bernd RosenowAAAI 2025 · 被引用 3 次
- Noise-Robust Learning from Multiple Unsupervised Sources of Inferred LabelsAmila Silva, Ling Luo, Shanika Karunasekera, Christopher LeckieAAAI 2022 · 被引用 11 次
- Multi-label Iterated Learning for Image Classification with Label AmbiguitySai Rajeswar, Pau Rodríguez, Soumye Singhal, David Vázquez 等CVPR 2022
