On the Modularity of Hypernetworks
Tomer Galanti, Lior Wolf
Abstract
In the context of learning to map an input to a function , two alternative methods are compared: (i) an embedding-based method, which learns a fixed function in which is encoded as a conditioning signal and the learned function takes the form , and (ii) hypernetworks, in which the weights of the function are given by a hypernetwork as . In this paper, we define the property of modularity as the ability to effectively learn a different function for each input instance . For this purpose, we adopt an expressivity perspective of this property and extend the theory of Devore et al. 1996 and provide a lower bound on the complexity (number of trainable parameters) of neural networks as function approximators, by eliminating the requirements for the approximation method to be robust. Our results are then used to compare the complexities of and , showing that under certain conditions and when letting the functions and be as large as we wish, can be smaller than by orders of magnitude. This sheds light on the modularity of hypernetworks in comparison with the embedding-based method. Besides, we show that for a structured target function, the overall number of trainable parameters in a hypernetwork is smaller by orders of magnitude than the number of trainable parameters of a standard neural network and an embedding method.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 43acd254-4402-4ddc-b7ed-bfe5f74972b8Cited by top-tier papers19
- Linear Transformers Are Secretly Fast Weight ProgrammersImanol Schlag, Kazuki Irie, Jürgen SchmidhuberICML 2021 · 394 citations
- Recomposing the Reinforcement Learning Building Blocks with HypernetworksElad Sarafian, Shai Keynan, Sarit KrausICML 2021 · 42 citations
- Universal Morphology Control via Contextual ModulationZheng Xiong, Jacob Beck, Shimon WhitesonICML 2023 · 27 citations
- Hypernetworks for Zero-Shot Transfer in Reinforcement LearningSahand Rezaei-Shoshtari, Charlotte Morissette, François Robert Hogan, Gregory Dudek et al.AAAI 2023 · 23 citations
- On Infinite-Width HypernetworksEtai Littwin, Tomer Galanti, Lior Wolf, Greg YangNeurIPS 2020 · 21 citations
Builds on4
- Continual learning with hypernetworksJohannes von Oswald, Christian Henning, João Sacramento, Benjamin F. GreweICLR 2020 · 412 citations
- Multiplicative Interactions and Where to Find ThemSiddhant M. Jayakumar, Wojciech M. Czarnecki, Jacob Menick, Jonathan Schwarz et al.ICLR 2020 · 152 citations
- Deep Meta Functionals for Shape RepresentationGidi Littwin, Lior WolfICCV 2019 · 90 citations
- Principled Weight Initialization for HypernetworksOscar Chang, Lampros Flokas, Hod LipsonICLR 2020 · 87 citations
Related papers
- Improving Memory Efficiency for Training KANs via Meta LearningZhangchi Zhao, Jun Shu, Deyu Meng, Zongben XuICML 2025
- HyperDeepONet: learning operator with complex target function space using the limited resources via hypernetworkJae Yong Lee, Sung Woong Cho, Hyung Ju HwangICLR 2023 · 3 citations
- Rational Neural Networks have Expressivity AdvantagesMaosen Tang, Alex TownsendICML 2026 · 1 citation
- The Expressive Power of Low-Rank AdaptationYuchen Zeng, Kangwook LeeICLR 2024 · 116 citations
- On the Expressive Power of Permutation-Equivariant Weight-Space NetworksAdir Dayan, Yam Eitan, Haggai MaronICML 2026
