Floating-Point Networks with Automatic Differentiation Can Represent Almost All Floating-Point Functions and Their Gradients
Sejun Park, Yeachan Park, Geonho Hwang
摘要
Theoretical studies show that for any differentiable function on a compact domain, there exists a neural network that approximates both the function values and gradients. However, such a result cannot be used in practice since it assumes real parameters and exact internal operations. In contrast, real implementations only use a finite subset of reals and machine operations with round-off errors. In this work, we investigate whether a similar result holds for neural networks under floating-point arithmetic, when the gradient with respect to the input is computed by the automatic differentiation algorithm . We first show that given a floating-point function (e.g., a loss function), arbitrary function values and gradients can be represented by a floating-point network and , respectively. We further extend this result: given , can simultaneously represent arbitrary gradients while represents the target values, under mild conditions. Our results hold for practical activation functions, e.g., ReLU, ELU, GELU, Swish, Sigmoid, and tanh.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper4
- Are Transformers universal approximators of sequence-to-sequence functions?Chulhee Yun, Srinadh Bhojanapalli, Ankit Singh Rawat, Sashank J. Reddi 等ICLR 2020 · 被引用 481 次
- Optimal Minimum Width for the Universal Approximation of Continuously Differentiable Functions by Deep Narrow MLPsGeonho HwangNeurIPS 2025 · 被引用 2 次
- Floating-Point Neural Networks Can Represent Almost All Floating-Point FunctionsGeonho Hwang, Yeachan Park, Wonyeol Lee, Sejun ParkICML 2025
- Floating-Point Neural Networks are Provably Robust Universal ApproximatorsGeonho Hwang, Wonyeol Lee, Yeachan Park, Sejun Park 等CAV 2025
相关 Paper
- On Minimum Depth and Width of Floating-Point Neural Networks for Representing Floating-Point FunctionsSejun Park, Yeachan Park, Geonho HwangICML 2026
- On the Correctness of Automatic Differentiation for Neural Networks with Machine-Representable ParametersWonyeol Lee, Sejun Park, Alex AikenICML 2023 · 被引用 6 次
- What does automatic differentiation compute for neural networks?Sejun Park, Sanghyuk Chun, Wonyeol LeeICLR 2024
- On Correctness of Automatic Differentiation for Non-Differentiable FunctionsWonyeol Lee, Hangyeol Yu, Xavier Rival, Hongseok YangNeurIPS 2020 · 被引用 50 次
- Faster Activation Functions at the Edge for Post-Training SpeedupsAnton Lydike, Jun Bi, Jackson WoodruffICML 2026
