Nonparametric Classification on Low Dimensional Manifolds using Overparameterized Convolutional Residual Networks
Zixuan Zhang, Kaiqi Zhang, Minshuo Chen, Yuma Takeda, Mengdi Wang, Tuo Zhao, Yu-Xiang Wang
摘要
Convolutional residual neural networks (ConvResNets), though overparameterized, can achieve remarkable prediction performance in practice, which cannot be well explained by conventional wisdom. To bridge this gap, we study the performance of ConvResNeXts, which cover ConvResNets as a special case, trained with weight decay from the perspective of nonparametric classification. Our analysis allows for infinitely many building blocks in ConvResNeXts, and shows that weight decay implicitly enforces sparsity on these blocks. Specifically, we consider a smooth target function supported on a low-dimensional manifold, then prove that ConvResNeXts can adapt to the function smoothness and low-dimensional structures and efficiently learn the function without suffering from the curse of dimensionality. Our findings partially justify the advantage of overparameterized ConvResNeXts over conventional machine learning models.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper4
- Besov Function Approximation and Binary Classification on Low-Dimensional Manifolds Using Convolutional Residual NetworksHao Liu, Minshuo Chen, Tuo Zhao, Wenjing LiaoICML 2021 · 被引用 42 次
- Provable Guarantees for Nonlinear Feature Learning in Three-Layer Neural NetworksEshaan Nichani, Alex Damian, Jason D. LeeNeurIPS 2023 · 被引用 24 次
- Learning Hierarchical Polynomials with Three-Layer Neural NetworksZihao Wang, Eshaan Nichani, Jason D. LeeICLR 2024 · 被引用 7 次
- Deep Learning meets Nonparametric Regression: Are Weight-Decayed DNNs Locally Adaptive?Kaiqi Zhang, Yu-Xiang WangICLR 2023 · 被引用 3 次
相关 Paper
- Benefits of Overparameterized Convolutional Residual Networks: Function Approximation under Smoothness ConstraintHao Liu, Minshuo Chen, Siawpeng Er, Wenjing Liao 等ICML 2022 · 被引用 16 次
- Approximation with CNNs in Sobolev Space: with Applications to ClassificationGuohao Shen, Yuling Jiao, Yuanyuan Lin, Jian HuangNeurIPS 2022 · 被引用 25 次
- Random Sparse Lifts: Construction, Analysis and Convergence of finite sparse networksDavid A. R. Robin, Kevin Scaman, Marc LelargeICLR 2024
- On the Expressive Power of Mixture-of-Experts for Structured Complex TasksMingze Wang, Weinan ENeurIPS 2025 · 被引用 3 次
- When Do Neural Networks Outperform Kernel Methods?Behrooz Ghorbani, Song Mei, Theodor Misiakiewicz, Andrea MontanariNeurIPS 2020 · 被引用 217 次
