Sampling weights of deep neural networks
Erik Lien Bolager, Iryna Burak, Chinmay Datar, Qing Sun, Felix Dietrich
摘要
We introduce a probability distribution, combined with an efficient sampling algorithm, for weights and biases of fully-connected neural networks. In a supervised learning context, no iterative optimization or gradient computations of internal network parameters are needed to obtain a trained network. The sampling is based on the idea of random feature models. However, instead of a data-agnostic distribution, e.g., a normal distribution, we use both the input and the output training data to sample shallow and deep networks. We prove that sampled networks are universal approximators. For Barron functions, we show that the -approximation error of sampled shallow networks decreases with the square root of the number of neurons. Our sampling scheme is invariant to rigid body transformations and scaling of the input data, which implies many popular pre-processing techniques are not required. In numerical experiments, we demonstrate that sampled networks achieve accuracy comparable to iteratively trained ones, but can be constructed orders of magnitude faster. Our test cases involve a classification benchmark from OpenML, sampling of neural operators to represent maps in function spaces, and transfer learning using well-known architectures.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Fast training of accurate physics-informed neural networks without gradient descentChinmay Datar, Taniya Kapoor, Abhishek Chandra, Qing Sun 等ICLR 2026 · 被引用 10 次
- DeepAFL: Deep Analytic Federated LearningJianheng Tang, Yajiang Huang, Kejia Fan, Feijiang Han 等ICLR 2026 · 被引用 5 次
- Rapid Training of Hamiltonian Graph Networks Using Random FeaturesAtamert Rahma, Chinmay Datar, Ana Cukarska, Felix DietrichICLR 2026 · 被引用 2 次
- Random Feature Representation BoostingNikita Zozoulenko, Thomas Cass, Lukas GononICML 2025
它引用的顶会 Paper4
- Fourier Neural Operator for Parametric Partial Differential EquationsZongyi Li, Nikola Borislavov Kovachki, Kamyar Azizzadenesheli, Burigede Liu 等ICLR 2021 · 被引用 3,911 次
- Hyper-Representations as Generative Models: Sampling Unseen Neural Network WeightsKonstantin Schürholt, Boris Knyazev, Xavier Giró-i-Nieto, Damian BorthNeurIPS 2022 · 被引用 78 次
- On the Existence of Universal Lottery TicketsRebekka Burkholz, Nilanjana Laha, Rajarshi Mukherjee, Alkis GotovosICLR 2022 · 被引用 38 次
- Transform Once: Efficient Operator Learning in Frequency DomainMichael Poli, Stefano Massaroli, Federico Berto, Jinkyoo Park 等NeurIPS 2022 · 被引用 29 次
相关 Paper
- Deep Ridgelet Transform and Unified Universality Theorem for Deep and Shallow Joint-Group-Equivariant MachinesSho Sonoda, Yuka Hashimoto, Isao Ishikawa, Masahiro IkedaICML 2025
- Batch normalization is sufficient for universal function approximation in CNNsRebekka BurkholzICLR 2024 · 被引用 8 次
- Expressive probabilistic sampling in recurrent neural networksShirui Chen, Linxing Jiang, Rajesh P. N. Rao, Eric Shea-BrownNeurIPS 2023 · 被引用 4 次
- Neural tangent kernels, transportation mappings, and universal approximationZiwei Ji, Matus Telgarsky, Ruicheng XianICLR 2020 · 被引用 45 次
- Scalable Neural Network KernelsArijit Sehanobish, Krzysztof Marcin Choromanski, Yunfan Zhao, Kumar Avinava Dubey 等ICLR 2024 · 被引用 9 次
