Leveraging Inter-Rater Agreement for Classification in the Presence of Noisy Labels
Maria Sofia Bucarelli, Lucas Cassano, Federico Siciliano, Amin Mantrach, Fabrizio Silvestri
摘要
In practical settings, classification datasets are obtained through a labelling process that is usually done by humans. Labels can be noisy as they are obtained by aggregating the different individual labels assigned to the same sample by multiple, and possibly disagreeing, annotators. The interrater agreement on these datasets can be measured while the underlying noise distribution to which the labels are subject is assumed to be unknown. In this work, we: (i) show how to leverage the inter-annotator statistics to estimate the noise distribution to which labels are subject; (ii) introduce methods that use the estimate of the noise distribution to learn from the noisy dataset; and (iii) establish generalization bounds in the empirical risk minimization framework that depend on the estimated quantities. We conclude the paper by providing experiments that illustrate our findings.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Learning from Noisy Labels via Conditional Distributionally Robust OptimizationHui Guo, Grace Y. Yi, Boyu WangNeurIPS 2024 · 被引用 8 次
- Visual Objectification in Films: Towards a New AI Task for Video InterpretationJulie Tores, Lucile Sassatelli, Hui-Yin Wu, Clement Bergman 等CVPR 2024 · 被引用 3 次
- On Group Sufficiency Under Label BiasHaoran Zhang, Olawale Salaudeen, Marzyeh GhassemiNeurIPS 2025 · 被引用 2 次
它引用的顶会 Paper8
- Symmetric Cross Entropy for Robust Learning With Noisy LabelsYisen Wang, Xingjun Ma, Zaiyi Chen, Yuan Luo 等ICCV 2019 · 被引用 1,125 次
- Human Uncertainty Makes Classification More RobustJoshua C. Peterson, Ruairidh M. Battleday, Thomas L. Griffiths, Olga RussakovskyICCV 2019 · 被引用 362 次
- Dual T: Reducing Estimation Error for Transition Matrix in Label-noise LearningYu Yao, Tongliang Liu, Bo Han, Mingming Gong 等NeurIPS 2020 · 被引用 297 次
- Can gradient clipping mitigate label noise?Aditya Krishna Menon, Ankit Singh Rawat, Sashank J. Reddi, Sanjiv KumarICLR 2020 · 被引用 163 次
- Clusterability as an Alternative to Anchor Points When Learning with Noisy LabelsZhaowei Zhu, Yiwen Song, Yang LiuICML 2021 · 被引用 112 次
相关 Paper
- To Aggregate or Not? Learning with Separate Noisy LabelsJiaheng Wei, Zhaowei Zhu, Tianyi Luo, Ehsan Amid 等KDD 2023 · 被引用 21 次
- Peer Loss Functions: Learning from Noisy Labels without Knowing Noise RatesYang Liu, Hongyi GuoICML 2020 · 被引用 280 次
- Adversarial Multi Class Learning under Weak Supervision with Performance GuaranteesAlessio Mazzetto, Cyrus Cousins, Dylan Sam, Stephen H. Bach 等ICML 2021 · 被引用 39 次
- Confidence Difference Reflects Various Supervised Signals in Confidence-Difference ClassificationYuanchao Dai, Ximing Li, Changchun LiICML 2025
- RATT: Leveraging Unlabeled Data to Guarantee GeneralizationSaurabh Garg, Sivaraman Balakrishnan, J. Zico Kolter, Zachary C. LiptonICML 2021 · 被引用 30 次
