Reproducing Kernel Banach Space Models for Neural Networks with Application to Rademacher Complexity Analysis
Alistair Shilton, Sunil Gupta, Santu Rana, Svetha Venkatesh
摘要
This paper explores the use of Hermite transform based reproducing kernel Banach space methods to construct exact or un-approximated models of feedforward neural networks of arbitrary width, depth and topology, including ResNet and Transformers networks, assuming only a feedforward topology, finite energy activations and finite (spectral-) norm weights and biases. Using this model, two straightforward but surprisingly tight bounds on Rademacher complexity are derived, precisely (1) a general bound that is width-independent and scales exponentially with depth; and
(2) a width-and depth-independent bound for networks with appropriately constrained (below threshold) weights and biases.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper2
- Beyond Linearization: On Quadratic and Higher-Order Approximation of Wide Neural NetworksYu Bai, Jason D. LeeICLR 2020 · 被引用 128 次
- Gradient Descent in Neural Networks as Sequential Learning in Reproducing Kernel Banach SpaceAlistair Shilton, Sunil Gupta, Santu Rana, Svetha VenkateshICML 2023 · 被引用 3 次
相关 Paper
- Why High-rank Neural Networks Generalize?: An Algebraic Framework with RKHSsYuka Hashimoto, Sho Sonoda, Isao Ishikawa, Masahiro IkedaICLR 2026 · 被引用 1 次
- On Measuring Excess Capacity in Neural NetworksFlorian Graf, Sebastian Zeng, Bastian Rieck, Marc Niethammer 等NeurIPS 2022 · 被引用 13 次
- Characterizing the spectrum of the NTK via a power series expansionMichael Murray, Hui Jin, Benjamin Bowman, Guido MontúfarICLR 2023 · 被引用 2 次
- Tighter Sparse Approximation Bounds for ReLU Neural NetworksCarles Domingo-Enrich, Youssef MrouehICLR 2022 · 被引用 4 次
- Why Do Deep Residual Networks Generalize Better than Deep Feedforward Networks? - A Neural Tangent Kernel PerspectiveKaixuan Huang, Yuqing Wang, Molei Tao, Tuo ZhaoNeurIPS 2020 · 被引用 107 次
