Lune

NeurIPS2020顶会

Sharp Representation Theorems for ReLU Networks with Precise Dependence on Depth

Guy Bresler, Dheeraj Nagaraj

2020年份
27被引次数
4顶会引用

摘要

We prove sharp dimension-free representation results for neural networks with DD ReLU layers under square loss for a class of functions GD\mathcal{G}_D defined in the paper. These results capture the precise benefits of depth in the following sense:

  1. The rates for representing the class of functions GD\mathcal{G}_D via DD ReLU layers is sharp up to constants, as shown by matching lower bounds.
  2. For each DD, GD⊆GD+1\mathcal{G}_{D} \subseteq \mathcal{G}_{D+1} and as DD grows the class of functions GD\mathcal{G}_{D} contains progressively less smooth functions.
  3. If D′<DD^{\prime} < D, then the approximation rate for the class GD\mathcal{G}_D achieved by depth D′D^{\prime} networks is strictly worse than that achieved by depth DD networks. This constitutes a fine-grained characterization of the representation power of feedforward networks of arbitrary depth DD and number of neurons NN, in contrast to existing representation results which either require DD growing quickly with NN or assume that the function being represented is highly smooth. In the latter case similar rates can be obtained with a single nonlinear layer. Our results confirm the prevailing hypothesis that deeper networks are better at representing less smooth functions, and indeed, the main technical novelty is to fully exploit the fact that deep networks can produce highly oscillatory functions with few activation functions.

问问这篇 Paper

智能体会读完全文。

Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。

可以从这些问题问起

智能体调用

Luneget_paper_fulltext

在 Lune 里问

免费开始,无需绑卡

引用它的顶会 Paper4

问问它们各自怎么用它

它引用的顶会 Paper2

相关 Paper

黄昏的海面,两侧是细线勾勒的悬崖