SuperLoss: A Generic Loss for Robust Curriculum Learning
Thibault Castells, Philippe Weinzaepfel, Jérôme Revaud
摘要
Curriculum learning is a technique to improve a model performance and generalization based on the idea that easy samples should be presented before difficult ones during training. While it is generally complex to estimate a priori the difficulty of a given sample, recent works have shown that curriculum learning can be formulated dynamically in a self-supervised manner. The key idea is to somehow estimate the importance (or weight) of each sample directly during training based on the observation that easy and hard samples behave differently and can therefore be separated. However, these approaches are usually limited to a specific task (e.g., classification) and require extra data annotations, layers or parameters as well as a dedicated training procedure. We propose instead a simple and generic method that can be applied to a variety of losses and tasks without any change in the learning procedure. It consists in appending a novel loss function on top of any existing task loss, hence its name: the SuperLoss. Its main effect is to automatically downweight the contribution of samples with a large loss, i.e. hard samples, effectively mimicking the core principle of curriculum learning. As a side effect, we show that our loss prevents the memorization of noisy samples, making it possible to train from noisy data even with non-robust loss functions. Experimental results on image classification, regression, object detection and image retrieval demonstrate consistent gain, particularly in the presence of noise.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper16
- Adap-τ : Adaptively Modulating Embedding Magnitude for RecommendationJiawei Chen, Junkang Wu, Jiancan Wu, Xuezhi Cao 等WWW 2023 · 被引用 48 次
- Confusion-Resistant Federated Learning via Diffusion-Based Data Harmonization on Non-IID DataXiaohong Chen, Canran Xiao, Yongmei LiuNeurIPS 2024 · 被引用 41 次
- Curriculum Co-disentangled Representation Learning across Multiple Environments for Social RecommendationXin Wang, Zirui Pan, Yuwei Zhou, Hong Chen 等ICML 2023 · 被引用 31 次
- Intra- and Inter-Modal Curriculum for Multimodal LearningYuwei Zhou, Xin Wang, Hong Chen, Xuguang Duan 等ACM MM 2023 · 被引用 28 次
- CurBench: Curriculum Learning BenchmarkYuwei Zhou, Zirui Pan, Xin Wang, Hong Chen 等ICML 2024 · 被引用 11 次
它引用的顶会 Paper10
- DivideMix: Learning with Noisy Labels as Semi-supervised LearningJunnan Li, Richard Socher, Steven C. H. HoiICLR 2020 · 被引用 1,326 次
- SELF: Learning to Filter Noisy Labels with Self-EnsemblingDuc Tam Nguyen, Chaithanya Kumar Mummadi, Thi-Phuong-Nhung Ngo, Thi Hoai Phuong Nguyen 等ICLR 2020 · 被引用 354 次
- NLNL: Negative Learning for Noisy LabelsYoungdong Kim, Junho Yim, Juseung Yun, Junmo KimICCV 2019 · 被引用 338 次
- Deep Self-Learning From Noisy LabelsJiangfan Han, Ping Luo, Xiaogang WangICCV 2019 · 被引用 315 次
- O2U-Net: A Simple Noisy Label Detection Approach for Deep Neural NetworksJinchi Huang, Lie Qu, Rongfei Jia, Binqiang ZhaoICCV 2019 · 被引用 276 次
相关 Paper
- Robust Curriculum Learning: from clean label detection to noisy label self-correctionTianyi Zhou, Shengjie Wang, Jeff A. BilmesICLR 2021 · 被引用 111 次
- Curriculum Loss: Robust Learning and Generalization against Label CorruptionYueming Lyu, Ivor W. TsangICLR 2020 · 被引用 190 次
- CurricularFace: Adaptive Curriculum Learning Loss for Deep Face RecognitionYuge Huang, Yuhan Wang, Ying Tai, Xiaoming Liu 等CVPR 2020
- Mitigating Label Noise through Data AmbiguationJulian Lienen, Eyke HüllermeierAAAI 2024 · 被引用 14 次
- Curriculum Learning Meets Weakly Supervised Multimodal Correlation LearningSijie Mai, Ya Sun, Haifeng HuEMNLP 2022 · 被引用 9 次
