Lookahead: A Far-sighted Alternative of Magnitude-based Pruning
Sejun Park, Jaeho Lee, Sangwoo Mo, Jinwoo Shin
Abstract
Magnitude-based pruning is one of the simplest methods for pruning neural networks. Despite its simplicity, magnitude-based pruning and its variants demonstrated remarkable performances for pruning modern architectures. Based on the observation that the magnitude-based pruning indeed minimizes the Frobenius distortion of a linear operator corresponding to a single layer, we develop a simple pruning method, coined lookahead pruning, by extending the single layer optimization to a multi-layer optimization. Our experimental results demonstrate that the proposed method consistently outperforms the magnitude pruning on various networks including VGG and ResNet, particularly in the high-sparsity regime.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers20
- Pruning neural networks without any data by iteratively conserving synaptic flowHidenori Tanaka, Daniel Kunin, Daniel L. K. Yamins, Surya GanguliNeurIPS 2020 · 884 citations
- Layer-adaptive Sparsity for the Magnitude-based PruningJaeho Lee, Sejun Park, Sangwoo Mo, Sungsoo Ahn et al.ICLR 2021 · 331 citations
- Structural Pruning for Diffusion ModelsGongfan Fang, Xinyin Ma, Xinchao WangNeurIPS 2023 · 257 citations
- GDP: Stabilized Neural Network Pruning via Gates with Differentiable PolarizationYi Guo, Huan Yuan, Jianchao Tan, Zhangyang Wang et al.ICCV 2021 · 52 citations
- Addressing Loss of Plasticity and Catastrophic Forgetting in Continual LearningMohamed Elsayed, A. Rupam MahmoodICLR 2024 · 52 citations
Builds on1
Related papers
- Greedy Optimization Provably Wins the Lottery: Logarithmic Number of Winning Tickets is EnoughMao Ye, Lemeng Wu, Qiang LiuNeurIPS 2020 · 17 citations
- Efficient Joint Optimization of Layer-Adaptive Weight Pruning in Deep Neural NetworksKaixin Xu, Zhe Wang, Xue Geng, Min Wu et al.ICCV 2023 · 24 citations
- Layer-wise Model Pruning based on Mutual InformationChun Fan, Jiwei Li, Tianwei Zhang, Xiang Ao et al.EMNLP 2021
- Movement Pruning: Adaptive Sparsity by Fine-TuningVictor Sanh, Thomas Wolf, Alexander M. RushNeurIPS 2020 · 656 citations
- What Makes a Good Prune? Maximal Unstructured Pruning for Maximal Cosine SimilarityGabryel Mason-Williams, Fredrik DahlqvistICLR 2024 · 20 citations
