Lune

NeurIPS2023Top-tier venue

Most Neural Networks Are Almost Learnable

Amit Daniely, Nati Srebro, Gal Vardi

2023Year
1Citations
1Top-tier citations

Abstract

We present a PTAS for learning random constant-depth networks. We show that for any fixed ϵ>0\epsilon>0 and depth ii, there is a poly-time algorithm that for any distribution on d⋅Sd−1\sqrt{d} \cdot \mathbb{S}^{d-1} learns random Xavier networks of depth ii, up to an additive error of ϵ\epsilon. The algorithm runs in time and sample complexity of (dˉ)poly(ϵ−1)(\bar{d})^{\mathrm{poly}(\epsilon^{-1})}, where dˉ\bar d is the size of the network. For some cases of sigmoid and ReLU-like activations the bound can be improved to (dˉ)polylog(ϵ−1)(\bar{d})^{\mathrm{polylog}(\epsilon^{-1})}, resulting in a quasi-poly-time algorithm for learning constant depth random networks.

Ask about this paper

Your agent reads all of it.

Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.

Questions to start from

Your agent calls

Luneget_paper_fulltext

Ask in Lune

Free to start. No credit card required.

Cited by top-tier papers1

Ask how each one uses it

Builds on6

Related papers

Dusk over the sea between two cliffs drawn in fine vertical lines