Rational neural networks
Nicolas Boullé, Yuji Nakatsukasa, Alex Townsend
2020Year
130Citations
14Top-tier citations
Abstract
We consider neural networks with rational activation functions. The choice of the nonlinear activation function in deep learning architectures is crucial and heavily impacts the performance of a neural network. We establish optimal bounds in terms of network complexity and prove that rational neural networks approximate smooth functions more efficiently than ReLU networks with exponentially smaller depth. The flexibility and smoothness of rational activation functions make them an attractive alternative to ReLU, as we demonstrate with numerical experiments.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers14
- Convolutional Neural Operators for robust and accurate learning of PDEsBogdan Raonic, Roberto Molinaro, Tim De Ryck, Tobias Rohner et al.NeurIPS 2023 · 292 citations
- Diffusion Models are Minimax Optimal Distribution EstimatorsKazusato Oko, Shunta Akiyama, Taiji SuzukiICML 2023 · 152 citations
- Randomized Sparse Neural Galerkin Schemes for Solving Evolution Equations with Deep NetworksJules Berman, Benjamin PeherstorferNeurIPS 2023 · 42 citations
- A generalization of the randomized singular value decompositionNicolas Boullé, Alex TownsendICLR 2022 · 18 citations
- Learning on a Razor's Edge: Identifiability and Singularity of Polynomial Neural NetworksVahid Shahverdi, Giovanni Luca Marchetti, Kathlén KohnICLR 2026 · 11 citations
Builds on1
Related papers
- Rational Neural Networks have Expressivity AdvantagesMaosen Tang, Alex TownsendICML 2026 · 1 citation
- The phase diagram of approximation rates for deep neural networksDmitry Yarotsky, Anton ZhevnerchukNeurIPS 2020 · 156 citations
- Sharp Representation Theorems for ReLU Networks with Precise Dependence on DepthGuy Bresler, Dheeraj NagarajNeurIPS 2020 · 27 citations
- Shallow and Deep Networks are Near-Optimal Approximators of Korobov FunctionsMoïse Blanchard, Mohammed Amine BennounaICLR 2022 · 11 citations
- Compelling ReLU Networks to Exhibit Exponentially Many Linear Regions at Initialization and During TrainingMax Milkert, David Hyde, Forrest J. LaineICML 2025
