On the Modularity of Hypernetworks
Tomer Galanti, Lior Wolf
摘要
In the context of learning to map an input to a function , two alternative methods are compared: (i) an embedding-based method, which learns a fixed function in which is encoded as a conditioning signal and the learned function takes the form , and (ii) hypernetworks, in which the weights of the function are given by a hypernetwork as . In this paper, we define the property of modularity as the ability to effectively learn a different function for each input instance . For this purpose, we adopt an expressivity perspective of this property and extend the theory of Devore et al. 1996 and provide a lower bound on the complexity (number of trainable parameters) of neural networks as function approximators, by eliminating the requirements for the approximation method to be robust. Our results are then used to compare the complexities of and , showing that under certain conditions and when letting the functions and be as large as we wish, can be smaller than by orders of magnitude. This sheds light on the modularity of hypernetworks in comparison with the embedding-based method. Besides, we show that for a structured target function, the overall number of trainable parameters in a hypernetwork is smaller by orders of magnitude than the number of trainable parameters of a standard neural network and an embedding method.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper19
- Linear Transformers Are Secretly Fast Weight ProgrammersImanol Schlag, Kazuki Irie, Jürgen SchmidhuberICML 2021 · 被引用 394 次
- Recomposing the Reinforcement Learning Building Blocks with HypernetworksElad Sarafian, Shai Keynan, Sarit KrausICML 2021 · 被引用 42 次
- Universal Morphology Control via Contextual ModulationZheng Xiong, Jacob Beck, Shimon WhitesonICML 2023 · 被引用 27 次
- Hypernetworks for Zero-Shot Transfer in Reinforcement LearningSahand Rezaei-Shoshtari, Charlotte Morissette, François Robert Hogan, Gregory Dudek 等AAAI 2023 · 被引用 23 次
- On Infinite-Width HypernetworksEtai Littwin, Tomer Galanti, Lior Wolf, Greg YangNeurIPS 2020 · 被引用 21 次
它引用的顶会 Paper4
- Continual learning with hypernetworksJohannes von Oswald, Christian Henning, João Sacramento, Benjamin F. GreweICLR 2020 · 被引用 412 次
- Multiplicative Interactions and Where to Find ThemSiddhant M. Jayakumar, Wojciech M. Czarnecki, Jacob Menick, Jonathan Schwarz 等ICLR 2020 · 被引用 152 次
- Deep Meta Functionals for Shape RepresentationGidi Littwin, Lior WolfICCV 2019 · 被引用 90 次
- Principled Weight Initialization for HypernetworksOscar Chang, Lampros Flokas, Hod LipsonICLR 2020 · 被引用 87 次
相关 Paper
- Improving Memory Efficiency for Training KANs via Meta LearningZhangchi Zhao, Jun Shu, Deyu Meng, Zongben XuICML 2025
- HyperDeepONet: learning operator with complex target function space using the limited resources via hypernetworkJae Yong Lee, Sung Woong Cho, Hyung Ju HwangICLR 2023 · 被引用 3 次
- Rational Neural Networks have Expressivity AdvantagesMaosen Tang, Alex TownsendICML 2026 · 被引用 1 次
- The Expressive Power of Low-Rank AdaptationYuchen Zeng, Kangwook LeeICLR 2024 · 被引用 116 次
- On the Expressive Power of Permutation-Equivariant Weight-Space NetworksAdir Dayan, Yam Eitan, Haggai MaronICML 2026
