Lune

NeurIPS2023顶会

Learning in the Presence of Low-dimensional Structure: A Spiked Random Matrix Perspective

Jimmy Ba, Murat A. Erdogdu, Taiji Suzuki, Zhichao Wang, Denny Wu

2023年份
47被引次数
26顶会引用

摘要

We consider the problem of learning a single-index target function f * : R d → R under the spiked covariance data: where the link function σ * : R → R is a degree-p polynomial with information exponent k (defined as the lowest degree in the Hermite expansion of σ * ), and it depends on the projection of input x onto the spike (signal) direction µ ∈ R d . In the proportional asymptotic limit where the number of training examples n and the dimensionality d jointly diverge: n, d → ∞, n/d → ψ ∈ (0, ∞), we ask the following question: how large should the spike magnitude θ be, in order for (i) kernel methods, (ii) neural networks optimized by gradient descent, to learn f * ? We show that for kernel ridge regression, β ≥ 1 -1 p is both sufficient and necessary. Whereas for two-layer neural networks trained with gradient descent, β > 1 -1 k suffices. Our results demonstrate that both kernel methods and neural networks benefit from low-dimensional structures in the data. Further, since k ≤ p by definition, neural networks can adapt to such structures more effectively.

问问这篇 Paper

智能体会读完全文。

Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。

可以从这些问题问起

智能体调用

Luneget_paper_fulltext

在 Lune 里问

免费开始,无需绑卡

引用它的顶会 Paper26

问问它们各自怎么用它

它引用的顶会 Paper8

相关 Paper

黄昏的海面,两侧是细线勾勒的悬崖