Minimum Width for Universal Approximation using Squashable Activation Functions
Jonghyun Shin, Namjun Kim, Geonho Hwang, Sejun Park
摘要
The exact minimum width that allows for universal approximation of unbounded-depth networks is known only for ReLU and its variants. In this work, we study the minimum width of networks using general activation functions. Specifically, we focus on squashable functions that can approximate the identity function and binary step function by alternatively composing with affine transformations. We show that for networks using a squashable activation function to universally approximate L p functions from [0, 1] dx to R dy , the minimum width is maxdx, dy, 2 unless dx = dy = 1; the same bound holds for dx = dy = 1 if the activation function is monotone. We then provide sufficient conditions for squashability and show that all non-affine analytic functions and a class of piecewise functions are squashable, i.e., our minimum width result holds for those general classes of activation functions.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper4
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- Minimum Width for Universal ApproximationSejun Park, Chulhee Yun, Jaeho Lee, Jinwoo ShinICLR 2021 · 被引用 148 次
- Minimum width for universal approximation using ReLU networks on compact domainNamjun Kim, Chanho Min, Sejun ParkICLR 2024 · 被引用 19 次
- Achieve the Minimum Width of Neural Networks for Universal ApproximationYongqiang CaiICLR 2023 · 被引用 4 次
相关 Paper
- ReLU Network with Width d+O(1) Can Achieve Optimal Approximation RateChenghao Liu, Minghua ChenICML 2024 · 被引用 3 次
- Minimum Width for Deep, Narrow MLP: A Diffeomorphism ApproachGeonho HwangNeurIPS 2025 · 被引用 6 次
- Optimal Minimum Width for the Universal Approximation of Continuously Differentiable Functions by Deep Narrow MLPsGeonho HwangNeurIPS 2025 · 被引用 2 次
- The Implicit Bias of Minima Stability in Multivariate Shallow ReLU NetworksMor Shpigel Nacson, Rotem Mulayoff, Greg Ongie, Tomer Michaeli 等ICLR 2023 · 被引用 3 次
- Universal approximation power of deep residual neural networks via nonlinear control theoryPaulo Tabuada, Bahman GharesifardICLR 2021 · 被引用 31 次
