Lune

NeurIPS2022顶会

Neural Network Architecture Beyond Width and Depth

Shijun Zhang, Zuowei Shen, Haizhao Yang

2022年份
25被引次数
3顶会引用

摘要

This paper proposes a new neural network architecture by introducing an additional dimension called height beyond width and depth. Neural network architectures with height, width, and depth as hyper-parameters are called three-dimensional architectures. It is shown that neural networks with three-dimensional architectures are significantly more expressive than the ones with two-dimensional architectures (those with only width and depth as hyper-parameters), e.g., standard fully connected networks. The new network architecture is constructed recursively via a nested structure, and hence we call a network with the new architecture nested network (NestNet). A NestNet of height ss is built with each hidden neuron activated by a NestNet of height ≤s−1\le s-1. When s=1s=1, a NestNet degenerates to a standard network with a two-dimensional architecture. It is proved by construction that height-ss ReLU NestNets with O(n)\mathcal{O}(n) parameters can approximate 11-Lipschitz continuous functions on [0,1]d[0,1]^d with an error O(n−(s+1)/d)\mathcal{O}(n^{-(s+1)/d}), while the optimal approximation error of standard ReLU networks with O(n)\mathcal{O}(n) parameters is O(n−2/d)\mathcal{O}(n^{-2/d}). Furthermore, such a result is extended to generic continuous functions on [0,1]d[0,1]^d with the approximation error characterized by the modulus of continuity. Finally, we use numerical experimentation to show the advantages of the super-approximation power of ReLU NestNets.

问问这篇 Paper

智能体会读完全文。

Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。

可以从这些问题问起

智能体调用

Luneget_paper_fulltext

在 Lune 里问

免费开始,无需绑卡

引用它的顶会 Paper3

问问它们各自怎么用它

它引用的顶会 Paper6

相关 Paper

黄昏的海面,两侧是细线勾勒的悬崖