iFlood: A Stable and Effective Regularizer
Yuexiang Xie, Zhen Wang, Yaliang Li, Ce Zhang, Jingren Zhou, Bolin Ding
摘要
Various regularization methods have been designed to prevent overfitting of machine learning models. Among them, a surprisingly simple yet effective one, called Flooding, is proposed recently, which directly constrains the training loss on average to stay at a given level. However, our further studies uncover that the design of the loss function of Flooding can lead to a discrepancy between its objective and implementation, and cause the instability issue. To resolve these issues, in this paper, we propose a new regularizer, called individual Flood (denoted as iFlood). With instance-level constraints on training loss, iFlood encourages the trained models to better fit the under-fitted instances while suppressing the confidence on over-fitted ones. We theoretically show that the design of iFlood can be intrinsically connected with removing the noise or bias in training data, which makes it suitable for a variety of applications to improve the generalization performances of learned models. We also theoretically link iFlood to some other regularizers by comparing the inductive biases they introduce. Our experimental results on both image classification and language understanding tasks confirm that models learned with iFlood can stably converge to solutions with better generalization ability, and behave consistently at instance-level. By investigating how machine learning models learned with Flooding behave on the training instances and generalize on the testing ones, we uncover the instability issue of Flooding, where it can lead
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper6
- Sharpness-aware Minimization for Efficiently Improving GeneralizationPierre Foret, Ariel Kleiner, Hossein Mobahi, Behnam NeyshaburICLR 2021 · 被引用 1,861 次
- Do We Need Zero Training Loss After Achieving Zero Training Error?Takashi Ishida, Ikko Yamane, Tomoya Sakai, Gang Niu 等ICML 2020 · 被引用 155 次
- GradAug: A New Regularization Method for Deep Neural NetworksTaojiannan Yang, Sijie Zhu, Chen ChenNeurIPS 2020 · 被引用 43 次
- Generalized Entropy Regularization or: There's Nothing Special about Label SmoothingClara Meister, Elizabeth Salesky, Ryan CotterellACL 2020 · 被引用 4 次
- Regularizing Neural Networks via Adversarial Model PerturbationYaowei Zheng, Richong Zhang, Yongyi MaoCVPR 2021
相关 Paper
- Countering Overfitting with Counterfactual ExamplesFlavio Giorgi, Fabiano Veglianti, Fabrizio Silvestri, Gabriele TolomeiKDD 2026
- Understanding Gradient Regularization in Deep Learning: Efficient Finite-Difference Computation and Implicit BiasRyo Karakida, Tomoumi Takase, Tomohiro Hayase, Kazuki OsawaICML 2023 · 被引用 23 次
- Revisiting Learning with Noisy Labels: Active Forgetting and Noise SuppressionMengmeng Sheng, Zeren Sun, Tao Chen, Jinshan Pan 等CVPR 2026
- FedFixer: Mitigating Heterogeneous Label Noise in Federated LearningXinyuan Ji, Zhaowei Zhu, Wei Xi, Olga Gadyatskaya 等AAAI 2024 · 被引用 30 次
- Rethinking Out-of-Distribution Detection on Imbalanced Data DistributionKai Liu, Zhihang Fu, Sheng Jin, Chao Chen 等NeurIPS 2024 · 被引用 9 次
