MMA Regularization: Decorrelating Weights of Neural Networks by Maximizing the Minimal Angles
Zhennan Wang, Canqun Xiang, Wenbin Zou, Chen Xu
摘要
The strong correlation between neurons or filters can significantly weaken the generalization ability of neural networks. Inspired by the well-known Tammes problem, we propose a novel diversity regularization method to address this issue, which makes the normalized weight vectors of neurons or filters distributed on a hypersphere as uniformly as possible, through maximizing the minimal pairwise angles (MMA). This method can easily exert its effect by plugging the MMA regularization term into the loss function with negligible computational overhead. The MMA regularization is simple, efficient, and effective. Therefore, it can be used as a basic regularization method in neural network training. Extensive experiments demonstrate that MMA regularization is able to enhance the generalization ability of various modern models and achieves considerable performance improvements on CIFAR100 and TinyImageNet datasets. In addition, experiments on face verification show that MMA regularization is also effective for feature learning.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- Maximum Class Separation as Inductive Bias in One MatrixTejaswi Kasarla, Gertjan J. Burghouts, Max van Spengler, Elise van der Pol 等NeurIPS 2022 · 被引用 29 次
- Can We Evaluate Domain Adaptation Models Without Target-Domain Labels?Jianfei Yang, Hanjie Qian, Yuecong Xu, Kai Wang 等ICLR 2024 · 被引用 16 次
- FedLoGe: Joint Local and Generic Federated Learning under Long-tailed DataZikai Xiao, Zihan Chen, Liyinglan Liu, Yang Feng 等ICLR 2024 · 被引用 14 次
- Adversarial Parameter Attack on Deep Neural NetworksLijia Yu, Yihan Wang, Xiao-Shan GaoICML 2023 · 被引用 11 次
- Preventing Model Collapse in Deep Canonical Correlation Analysis by Noise RegularizationJunlin He, Jinxiao Du, Susu Xu, Wei MaNeurIPS 2024 · 被引用 5 次
它引用的顶会 Paper3
- Random Erasing Data AugmentationZhun Zhong, Liang Zheng, Guoliang Kang, Shaozi Li 等AAAI 2020 · 被引用 4,134 次
- Picking Winning Tickets Before Training by Preserving Gradient FlowChaoqi Wang, Guodong Zhang, Roger B. GrosseICLR 2020 · 被引用 743 次
- Regularizing Neural Networks via Minimizing Hyperspherical EnergyRongmei Lin, Weiyang Liu, Zhen Liu, Chen Feng 等CVPR 2020
相关 Paper
- Feature Variance Regularization: A Simple Way to Improve the Generalizability of Neural NetworksRanran Huang, Hanbo Sun, Ji Liu, Lu Tian 等AAAI 2020 · 被引用 5 次
- T-vMF Similarity for Regularizing Intra-Class Feature DistributionTakumi KobayashiCVPR 2021
- Orthogonal Over-Parameterized TrainingWeiyang Liu, Rongmei Lin, Zhen Liu, James M. Rehg 等CVPR 2021
- Neuron with Steady Response Leads to Better GeneralizationQiang Fu, Lun Du, Haitao Mao, Xu Chen 等NeurIPS 2022 · 被引用 5 次
- WLD-Reg: A Data-Dependent Within-Layer Diversity RegularizerFiras Laakom, Jenni Raitoharju, Alexandros Iosifidis, Moncef GabboujAAAI 2023 · 被引用 9 次
