Optimal Minimum Width for the Universal Approximation of Continuously Differentiable Functions by Deep Narrow MLPs
Geonho Hwang
摘要
In this paper, we investigate the universal approximation property of deep, narrow multilayer perceptrons (MLPs) for C 1 functions under the Sobolev norm, specifically the W 1 , ∞ norm. Although the optimal width of deep, narrow MLPs for approximating continuous functions has been extensively studied, significantly less is known about the corresponding optimal width for C 1 functions. We demonstrate that the optimal width can be determined in a wide range of cases within the C 1 setting. Our approach consists of two main steps. First, leveraging control theory, we show that any diffeomorphism can be approximated by deep, narrow MLPs. Second, using the Borsuk-Ulam theorem and various results from differential geometry, we prove that the optimal width for approximating arbitrary C 1 functions via diffeomorphisms is min( n + m, max(2 n + 1 , m )) in certain cases, including ( n, m ) = (8 , 8) and (16 , 8) , where n and m denote the input and output dimensions, respectively. Our results apply to a broad class of activation functions.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper4
- Minimum Width for Universal ApproximationSejun Park, Chulhee Yun, Jaeho Lee, Jinwoo ShinICLR 2021 · 被引用 148 次
- Minimum width for universal approximation using ReLU networks on compact domainNamjun Kim, Chanho Min, Sejun ParkICLR 2024 · 被引用 19 次
- Minimum Width for Deep, Narrow MLP: A Diffeomorphism ApproachGeonho HwangNeurIPS 2025 · 被引用 6 次
- Achieve the Minimum Width of Neural Networks for Universal ApproximationYongqiang CaiICLR 2023 · 被引用 4 次
相关 Paper
- Minimum Width of Leaky-ReLU Neural Networks for Uniform Universal ApproximationLi'ang Li, Yifei Duan, Guanghua Ji, Yongqiang CaiICML 2023 · 被引用 20 次
- Universal approximation power of deep residual neural networks via nonlinear control theoryPaulo Tabuada, Bahman GharesifardICLR 2021 · 被引用 31 次
- ReLU Network with Width d+O(1) Can Achieve Optimal Approximation RateChenghao Liu, Minghua ChenICML 2024 · 被引用 3 次
- Minimum Width for Universal Approximation using Squashable Activation FunctionsJonghyun Shin, Namjun Kim, Geonho Hwang, Sejun ParkICML 2025
- A closer look at the approximation capabilities of neural networksKai Fong Ernest ChongICLR 2020 · 被引用 18 次
