An Operator Theoretic View On Pruning Deep Neural Networks
William T. Redman, Maria Fonoberova, Ryan Mohr, Yannis G. Kevrekidis, Igor Mezic
摘要
The discovery of sparse subnetworks that are able to perform as well as full models has found broad applied and theoretical interest. While many pruning methods have been developed to this end, the naïve approach of removing parameters based on their magnitude has been found to be as robust as more complex, state-of-the-art algorithms. The lack of theory behind magnitude pruning's success, especially pre-convergence, and its relation to other pruning methods, such as gradient based pruning, are outstanding open questions in the field that are in need of being addressed. We make use of recent advances in dynamical systems theory, namely Koopman operator theory, to define a new class of theoretically motivated pruning algorithms. We show that these algorithms can be equivalent to magnitude and gradient based pruning, unifying these seemingly disparate methods, and find that they can be used to shed light on magnitude pruning's performance during the early part of training.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper11
- Generative Modeling of Regular and Irregular Time Series Data via Koopman VAEsIlan Naiman, N. Benjamin Erichson, Pu Ren, Michael W. Mahoney 等ICLR 2024 · 被引用 49 次
- Identifying Equivalent Training DynamicsWilliam T. Redman, Juan M. Bello-Rivas, Maria Fonoberova, Ryan Mohr 等NeurIPS 2024 · 被引用 15 次
- QuACK: Accelerating Gradient-Based Quantum Optimization with Koopman Operator LearningDi Luo, Jiayu Shen, Rumen Dangovski, Marin SoljacicNeurIPS 2023 · 被引用 11 次
- Koopman-based generalization bound: New aspect for full-rank weightsYuka Hashimoto, Sho Sonoda, Isao Ishikawa, Atsushi Nitanda 等ICLR 2024 · 被引用 6 次
- Enhancing Neural Training via a Correlated Dynamics ModelJonathan Brokman, Roy Betser, Rotem Turjeman, Tom Berkov 等ICLR 2024 · 被引用 5 次
它引用的顶会 Paper5
- Linear Mode Connectivity and the Lottery Ticket HypothesisJonathan Frankle, Gintare Karolina Dziugaite, Daniel M. Roy, Michael CarbinICML 2020 · 被引用 750 次
- Pruning Neural Networks at Initialization: Why Are We Missing the Mark?Jonathan Frankle, Gintare Karolina Dziugaite, Daniel M. Roy, Michael CarbinICLR 2021 · 被引用 261 次
- A Signal Propagation Perspective for Pruning Neural Networks at InitializationNamhoon Lee, Thalaiyasingam Ajanthan, Stephen Gould, Philip H. S. TorrICLR 2020 · 被引用 174 次
- Optimizing Neural Networks via Koopman Operator TheoryAkshunna S. Dogra, William T. RedmanNeurIPS 2020 · 被引用 65 次
- Validating the Lottery Ticket Hypothesis with Inertial Manifold TheoryZeru Zhang, Jiayin Jin, Zijie Zhang, Yang Zhou 等NeurIPS 2021 · 被引用 45 次
相关 Paper
- Optimal Recurrent Network Topologies for Dynamical Systems ReconstructionChristoph Jürgen Hemmer, Manuel Brenner, Florian Hess, Daniel DurstewitzICML 2024 · 被引用 6 次
- An Operator Theoretic Approach for Analyzing Sequence Neural NetworksIlan Naiman, Omri AzencotAAAI 2023 · 被引用 14 次
- A Gradient Flow Framework For Analyzing Network PruningEkdeep Singh Lubana, Robert P. DickICLR 2021 · 被引用 60 次
- What Makes a Good Prune? Maximal Unstructured Pruning for Maximal Cosine SimilarityGabryel Mason-Williams, Fredrik DahlqvistICLR 2024 · 被引用 20 次
- Lottery Tickets on a Data Diet: Finding Initializations with Sparse Trainable NetworksMansheej Paul, Brett W. Larsen, Surya Ganguli, Jonathan Frankle 等NeurIPS 2022 · 被引用 27 次
