Weight matrices compression based on PDB model in deep neural networks
Xiaoling Wu, Junpeng Zhu, Zeng Li
Abstract
Weight matrix compression has been demonstrated to effectively reduce overfitting and improve the generalization performance of deep neural networks. Compression is primarily achieved by filtering out noisy eigenvalues of the weight matrix. In this work, a novel Population Double Bulk (PDB) model is proposed to characterize the eigenvalue behavior of the weight matrix, which is more general than the existing Population Unit Bulk (PUB) model. Based on PDB model and Random Matrix Theory (RMT), we have discovered a new PDBLS algorithm for determining the boundary between noisy eigenvalues and information. A PDB Noise-Filtering algorithm is further introduced to reduce the rank of the weight matrix for compression. Experiments show that our PDB model fits the empirical distribution of eigenvalues of the weight matrix better than the PUB model, and our compressed weight matrices have lower rank at the same level of test accuracy. In some cases, our compression method can even improve generalization performance when labels contain noise. The code is avaliable at https://github.com/xlwu571/PDBLS .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on4
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Compressing Neural Networks: Towards Determining the Optimal Layer-wise DecompositionLucas Liebenwein, Alaa Maalouf, Dan Feldman, Daniela RusNeurIPS 2021 · 60 citations
- Low-Rank Compression of Neural Nets: Learning the Rank of Each LayerYerlan Idelbayev, Miguel Á. Carreira-PerpiñánCVPR 2020
Related papers
- PAC-Bayes Information BottleneckZifeng Wang, Shao-Lun Huang, Ercan Engin Kuruoglu, Jimeng Sun et al.ICLR 2022 · 42 citations
- Unifying Low Dimensional Spectra in Deep LearningConnall Garrod, Jonathan KeatingICML 2026 · 12 citations
- How does Weight Correlation Affect Generalisation Ability of Deep Neural Networks?Gaojie Jin, Xinping Yi, Liang Zhang, Lijun Zhang et al.NeurIPS 2020 · 5 citations
- Information-Theoretic Understanding of Population Risk Improvement with Model CompressionYuheng Bu, Weihao Gao, Shaofeng Zou, Venugopal V. VeeravalliAAAI 2020 · 18 citations
- Compression based bound for non-compressed network: unified generalization error analysis of large compressible deep neural networkTaiji Suzuki, Hiroshi Abe, Tomoaki NishimuraICLR 2020 · 57 citations
