Lune

NeurIPS2020顶会

On the Modularity of Hypernetworks

Tomer Galanti, Lior Wolf

2020年份
79被引次数
19顶会引用

摘要

In the context of learning to map an input II to a function hI:X→Rh_I:\mathcal{X}\to \mathbb{R}, two alternative methods are compared: (i) an embedding-based method, which learns a fixed function in which II is encoded as a conditioning signal e(I)e(I) and the learned function takes the form hI(x)=q(x,e(I))h_I(x) = q(x,e(I)), and (ii) hypernetworks, in which the weights θIθ_I of the function hI(x)=g(x;θI)h_I(x) = g(x;θ_I) are given by a hypernetwork ff as θI=f(I)θ_I=f(I). In this paper, we define the property of modularity as the ability to effectively learn a different function for each input instance II. For this purpose, we adopt an expressivity perspective of this property and extend the theory of Devore et al. 1996 and provide a lower bound on the complexity (number of trainable parameters) of neural networks as function approximators, by eliminating the requirements for the approximation method to be robust. Our results are then used to compare the complexities of qq and gg, showing that under certain conditions and when letting the functions ee and ff be as large as we wish, gg can be smaller than qq by orders of magnitude. This sheds light on the modularity of hypernetworks in comparison with the embedding-based method. Besides, we show that for a structured target function, the overall number of trainable parameters in a hypernetwork is smaller by orders of magnitude than the number of trainable parameters of a standard neural network and an embedding method.

问问这篇 Paper

智能体会读完全文。

Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。

可以从这些问题问起

智能体调用

Luneget_paper_fulltext

在 Lune 里问

免费开始,无需绑卡

引用它的顶会 Paper19

问问它们各自怎么用它

它引用的顶会 Paper4

相关 Paper

黄昏的海面,两侧是细线勾勒的悬崖