Minimum width for universal approximation using ReLU networks on compact domain
Namjun Kim, Chanho Min, Sejun Park
Abstract
It has been shown that deep neural networks of a large enough width are universal approximators but they are not if the width is too small. There were several attempts to characterize the minimum width enabling the universal approximation property; however, only a few of them found the exact values. In this work, we show that the minimum width for approximation of functions from to is exactly if an activation function is ReLU-Like (e.g., ReLU, GELU, Softplus). Compared to the known result for ReLU networks, when the domain is , our result first shows that approximation on a compact domain requires smaller width than on . We next prove a lower bound on for uniform approximation using general activation functions including ReLU: if . Together with our first result, this shows a dichotomy between and uniform approximations for general activation functions and input/output dimensions.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext eea9302c-7ba4-4876-a0c2-e9855f4fbe3fCited by top-tier papers10
- Minimum Width of Leaky-ReLU Neural Networks for Uniform Universal ApproximationLi'ang Li, Yifei Duan, Guanghua Ji, Yongqiang CaiICML 2023 · 20 citations
- Minimum Width for Deep, Narrow MLP: A Diffeomorphism ApproachGeonho HwangNeurIPS 2025 · 6 citations
- ReLU Network with Width d+O(1) Can Achieve Optimal Approximation RateChenghao Liu, Minghua ChenICML 2024 · 3 citations
- Optimal Minimum Width for the Universal Approximation of Continuously Differentiable Functions by Deep Narrow MLPsGeonho HwangNeurIPS 2025 · 2 citations
- Low-dimensional topology of deep neural networksJunyu Ren, Lek-Heng LimICML 2026
Builds on3
- Minimum Width for Universal ApproximationSejun Park, Chulhee Yun, Jaeho Lee, Jinwoo ShinICLR 2021 · 148 citations
- On the Optimal Memorization Power of ReLU Neural NetworksGal Vardi, Gilad Yehudai, Ohad ShamirICLR 2022 · 42 citations
- Achieve the Minimum Width of Neural Networks for Universal ApproximationYongqiang CaiICLR 2023 · 4 citations
Related papers
- Minimum Width for Universal Approximation using Squashable Activation FunctionsJonghyun Shin, Namjun Kim, Geonho Hwang, Sejun ParkICML 2025
- How Many Neurons Does it Take to Approximate the Maximum?Itay Safran, Daniel Reichman, Paul ValiantSODA 2024 · 3 citations
- On Minimum Depth and Width of Floating-Point Neural Networks for Representing Floating-Point FunctionsSejun Park, Yeachan Park, Geonho HwangICML 2026
- Shallow and Deep Networks are Near-Optimal Approximators of Korobov FunctionsMoïse Blanchard, Mohammed Amine BennounaICLR 2022 · 11 citations
- The phase diagram of approximation rates for deep neural networksDmitry Yarotsky, Anton ZhevnerchukNeurIPS 2020 · 156 citations
