How does Weight Correlation Affect Generalisation Ability of Deep Neural Networks?
Gaojie Jin, Xinping Yi, Liang Zhang, Lijun Zhang, Sven Schewe, Xiaowei Huang
摘要
This paper studies the novel concept of weight correlation in deep neural networks and discusses its impact on the networks' generalisation ability. For fully-connected layers, the weight correlation is defined as the average cosine similarity between weight vectors of neurons, and for convolutional layers, the weight correlation is defined as the cosine similarity between filter matrices. Theoretically, we show that, weight correlation can, and should, be incorporated into the PAC Bayesian framework for the generalisation of neural networks, and the resulting generalisation bound is monotonic with respect to the weight correlation. We formulate a new complexity measure, which lifts the PAC Bayes measure with weight correlation, and experimentally confirm that it is able to rank the generalisation errors of a set of networks more precisely than existing measures. More importantly, we develop a new regulariser for training, and provide extensive experiments that show that the generalisation error can be greatly reduced with our novel approach.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Data-IQ: Characterizing subgroups with heterogeneous outcomes in tabular dataNabeel Seedat, Jonathan Crabbé, Ioana Bica, Mihaela van der SchaarNeurIPS 2022 · 被引用 40 次
- Towards Green AI in Fine-tuning Large Language Models via Adaptive BackpropagationKai Huang, Hanyun Yin, Heng Huang, Wei GaoICLR 2024 · 被引用 21 次
- Confusion-Aware Spectral Regularizer for Long-Tailed RecognitionZiquan Zhu, Gaojie Jin, Hanruo Zhu, Si-Yuan Lu 等CVPR 2026 · 被引用 4 次
- TANGOS: Regularizing Tabular Neural Networks through Gradient Orthogonalization and SpecializationAlan Jeffares, Tennison Liu, Jonathan Crabbé, Fergus Imrie 等ICLR 2023 · 被引用 3 次
- Learning Verified Safe Neural Network Controllers for Multi-Agent Path FindingMingyue Zhang, Nianyu Li, Yi Chen, Jialong Li 等AAAI 2025 · 被引用 2 次
它引用的顶会 Paper3
- Fantastic Generalization Measures and Where to Find ThemYiding Jiang, Behnam Neyshabur, Hossein Mobahi, Dilip Krishnan 等ICLR 2020 · 被引用 705 次
- Generalization bounds for deep convolutional neural networksPhilip M. Long, Hanie SedghiICLR 2020 · 被引用 102 次
- The intriguing role of module criticality in the generalization of deep networksNiladri S. Chatterji, Behnam Neyshabur, Hanie SedghiICLR 2020 · 被引用 59 次
相关 Paper
- Neural Complexity MeasuresYoonho Lee, Juho Lee, Sung Ju Hwang, Eunho Yang 等NeurIPS 2020 · 被引用 13 次
- Singular Bayesian Neural NetworksMame Diarra Toure, David A StephensICML 2026 · 被引用 2 次
- On the Interpretability of Regularisation for Neural Networks Through Model Gradient SimilarityVincent Szolnoky, Viktor Andersson, Balázs Kulcsár, Rebecka JörnstenNeurIPS 2022 · 被引用 6 次
- Generalization Bounds for Rank-sparse Neural NetworksAntoine Ledent, Rodrigo Alves, Yunwen LeiNeurIPS 2025 · 被引用 4 次
- Koopman-based generalization bound: New aspect for full-rank weightsYuka Hashimoto, Sho Sonoda, Isao Ishikawa, Atsushi Nitanda 等ICLR 2024 · 被引用 6 次
