MMA Regularization: Decorrelating Weights of Neural Networks by Maximizing the Minimal Angles
Zhennan Wang, Canqun Xiang, Wenbin Zou, Chen Xu
Abstract
The strong correlation between neurons or filters can significantly weaken the generalization ability of neural networks. Inspired by the well-known Tammes problem, we propose a novel diversity regularization method to address this issue, which makes the normalized weight vectors of neurons or filters distributed on a hypersphere as uniformly as possible, through maximizing the minimal pairwise angles (MMA). This method can easily exert its effect by plugging the MMA regularization term into the loss function with negligible computational overhead. The MMA regularization is simple, efficient, and effective. Therefore, it can be used as a basic regularization method in neural network training. Extensive experiments demonstrate that MMA regularization is able to enhance the generalization ability of various modern models and achieves considerable performance improvements on CIFAR100 and TinyImageNet datasets. In addition, experiments on face verification show that MMA regularization is also effective for feature learning.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e5d9ff05-7bfd-41a9-9d15-7b7b80a9ae22Cited by top-tier papers7
- Maximum Class Separation as Inductive Bias in One MatrixTejaswi Kasarla, Gertjan J. Burghouts, Max van Spengler, Elise van der Pol et al.NeurIPS 2022 · 29 citations
- Can We Evaluate Domain Adaptation Models Without Target-Domain Labels?Jianfei Yang, Hanjie Qian, Yuecong Xu, Kai Wang et al.ICLR 2024 · 16 citations
- FedLoGe: Joint Local and Generic Federated Learning under Long-tailed DataZikai Xiao, Zihan Chen, Liyinglan Liu, Yang Feng et al.ICLR 2024 · 14 citations
- Adversarial Parameter Attack on Deep Neural NetworksLijia Yu, Yihan Wang, Xiao-Shan GaoICML 2023 · 11 citations
- Preventing Model Collapse in Deep Canonical Correlation Analysis by Noise RegularizationJunlin He, Jinxiao Du, Susu Xu, Wei MaNeurIPS 2024 · 5 citations
Builds on3
- Random Erasing Data AugmentationZhun Zhong, Liang Zheng, Guoliang Kang, Shaozi Li et al.AAAI 2020 · 4,134 citations
- Picking Winning Tickets Before Training by Preserving Gradient FlowChaoqi Wang, Guodong Zhang, Roger B. GrosseICLR 2020 · 743 citations
- Regularizing Neural Networks via Minimizing Hyperspherical EnergyRongmei Lin, Weiyang Liu, Zhen Liu, Chen Feng et al.CVPR 2020
Related papers
- Feature Variance Regularization: A Simple Way to Improve the Generalizability of Neural NetworksRanran Huang, Hanbo Sun, Ji Liu, Lu Tian et al.AAAI 2020 · 5 citations
- T-vMF Similarity for Regularizing Intra-Class Feature DistributionTakumi KobayashiCVPR 2021
- Orthogonal Over-Parameterized TrainingWeiyang Liu, Rongmei Lin, Zhen Liu, James M. Rehg et al.CVPR 2021
- Neuron with Steady Response Leads to Better GeneralizationQiang Fu, Lun Du, Haitao Mao, Xu Chen et al.NeurIPS 2022 · 5 citations
- WLD-Reg: A Data-Dependent Within-Layer Diversity RegularizerFiras Laakom, Jenni Raitoharju, Alexandros Iosifidis, Moncef GabboujAAAI 2023 · 9 citations
