Noise Attention Learning: Enhancing Noise Robustness by Gradient Scaling
Yangdi Lu, Yang Bo, Wenbo He
摘要
Machine learning has been highly successful in data-driven applications but is often hampered when the data contains noise, especially label noise. When trained on noisy labels, deep neural networks tend to fit all noisy labels, resulting in poor generalization. To handle this problem, a common idea is to force the model to fit only clean samples rather than mislabeled ones. In this paper, we propose a simple yet effective method that automatically distinguishes the mislabeled samples and prevents the model from memorizing them, named Noise Attention Learning. In our method, we introduce an attention branch to produce attention weights based on representations of samples. This attention branch is learned to divide the samples according to the predictive power in their representations. We design the corresponding loss function that incorporates the attention weights for training the model without affecting the original learning direction. Empirical results show that most of the mislabeled samples yield significantly lower weights than the clean ones. Furthermore, our theoretical analysis shows that the gradients of training samples are dynamically scaled by the attention weights, implicitly preventing memorization of the mislabeled samples. Experimental results on two benchmarks (CIFAR-10 and CIFAR-100) with simulated label noise and three realworld noisy datasets (ANIMAL-10N, Clothing1M and Webvision) demonstrate that our approach outperforms state-of-the-art methods.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Robust Classification via Regression for Learning with Noisy LabelsErik Englesson, Hossein AzizpourICLR 2024 · 被引用 12 次
- Theoretically Guaranteed Bidirectional Data Rectification for Robust Sequential RecommendationYatong Sun, Bin Wang, Zhu Sun, Xiaochun Yang 等NeurIPS 2023 · 被引用 11 次
- Subclass-Dominant Label Noise: A Counterexample for the Success of Early StoppingYingbin Bai, Zhongyi Han, Erkun Yang, Jun Yu 等NeurIPS 2023 · 被引用 10 次
- Revisiting Interpolation for Noisy Label CorrectionYuanzhuo Xu, Xiaoguang Niu, Jie Yang, Ruiyi Su 等AAAI 2025 · 被引用 8 次
- Handling Label Noise via Instance-Level Difficulty Modeling and Dynamic OptimizationKuan Zhang, Chengliang Chai, Jingzhe Xu, Chi Zhang 等NeurIPS 2025 · 被引用 6 次
它引用的顶会 Paper20
- RandAugment: Practical Automated Data Augmentation with a Reduced Search SpaceEkin Dogus Cubuk, Barret Zoph, Jonathon Shlens, Quoc LeNeurIPS 2020 · 被引用 4,453 次
- DivideMix: Learning with Noisy Labels as Semi-supervised LearningJunnan Li, Richard Socher, Steven C. H. HoiICLR 2020 · 被引用 1,326 次
- Symmetric Cross Entropy for Robust Learning With Noisy LabelsYisen Wang, Xingjun Ma, Zaiyi Chen, Yuan Luo 等ICCV 2019 · 被引用 1,125 次
- Early-Learning Regularization Prevents Memorization of Noisy LabelsSheng Liu, Jonathan Niles-Weed, Narges Razavian, Carlos Fernandez-GrandaNeurIPS 2020 · 被引用 798 次
- Normalized Loss Functions for Deep Learning with Noisy LabelsXingjun Ma, Hanxun Huang, Yisen Wang, Simone Romano 等ICML 2020 · 被引用 547 次
相关 Paper
- On the Role of Label Noise in the Feature Learning ProcessAndi Han, Wei Huang, Zhanpeng Zhou, Gang Niu 等ICML 2025
- Jo-SRC: A Contrastive Approach for Combating Noisy LabelsYazhou Yao, Zeren Sun, Chuanyi Zhang, Fumin Shen 等CVPR 2021
- DAT: Training Deep Networks Robust To Label-Noise by Matching the Feature DistributionsYuntao Qu, Shasha Mo, Jianwei NiuCVPR 2021
- Enhancing Robustness in Learning with Noisy Labels: An Asymmetric Co-Training ApproachMengmeng Sheng, Zeren Sun, Gensheng Pei, Tao Chen 等ACM MM 2024 · 被引用 7 次
- Large Loss Matters in Weakly Supervised Multi-Label ClassificationYoungwook Kim, Jae-Myung Kim, Zeynep Akata, Jungwoo LeeCVPR 2022 · 被引用 68 次
