Besov Function Approximation and Binary Classification on Low-Dimensional Manifolds Using Convolutional Residual Networks
Hao Liu, Minshuo Chen, Tuo Zhao, Wenjing Liao
Abstract
Most of existing statistical theories on deep neural networks have sample complexities cursed by the data dimension and therefore cannot well explain the empirical success of deep learning on high-dimensional data. To bridge this gap, we propose to exploit low-dimensional geometric structures of the real world data sets. We establish theoretical guarantees of convolutional residual networks (ConvResNet) in terms of function approximation and statistical estimation for binary classification. Specifically, given the data lying on a -dimensional manifold isometrically embedded in , we prove that if the network architecture is properly chosen, ConvResNets can (1) approximate Besov functions on manifolds with arbitrary accuracy, and (2) learn a classifier by minimizing the empirical logistic risk, which gives an excess risk in the order of , where is a smoothness parameter. This implies that the sample complexity depends on the intrinsic dimension , instead of the data dimension . Our results demonstrate that ConvResNets are adaptive to low-dimensional structures of data sets.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers11
- Understanding Scaling Laws with Statistical and Approximation Theory for Transformer Neural Networks on Intrinsically Low-dimensional DataAlexander Havrilla, Wenjing LiaoNeurIPS 2024 · 36 citations
- Hardness of Learning Neural Networks under the Manifold HypothesisBobak T. Kiani, Jason Wang, Melanie WeberNeurIPS 2024 · 25 citations
- Approximation with CNNs in Sobolev Space: with Applications to ClassificationGuohao Shen, Yuling Jiao, Yuanyuan Lin, Jian HuangNeurIPS 2022 · 25 citations
- Benefits of Overparameterized Convolutional Residual Networks: Function Approximation under Smoothness ConstraintHao Liu, Minshuo Chen, Siawpeng Er, Wenjing Liao et al.ICML 2022 · 16 citations
- Stable Minima Cannot Overfit in Univariate ReLU Networks: Generalization by Large Step SizesDan Qiao, Kaiqi Zhang, Esha Singh, Daniel Soudry et al.NeurIPS 2024 · 15 citations
Builds on1
Related papers
- Nonparametric Classification on Low Dimensional Manifolds using Overparameterized Convolutional Residual NetworksZixuan Zhang, Kaiqi Zhang, Minshuo Chen, Yuma Takeda et al.NeurIPS 2024 · 6 citations
- On Deep Generative Models for Approximation and Estimation of Distributions on ManifoldsBiraj Dahal, Alexander Havrilla, Minshuo Chen, Tuo Zhao et al.NeurIPS 2022 · 17 citations
- Deep Networks Provably Classify Data on CurvesTingran Wang, Sam Buchanan, Dar Gilboa, John WrightNeurIPS 2021 · 9 citations
- Effective Minkowski Dimension of Deep Nonparametric Regression: Function Approximation and Statistical TheoriesZixuan Zhang, Minshuo Chen, Mengdi Wang, Wenjing Liao et al.ICML 2023 · 4 citations
- Blessing of Dimensionality for Approximating Sobolev Classes on ManifoldsHong Ye Tan, Subhadip Mukherjee, Junqi Tang, Carola-Bibiane SchönliebAAAI 2026
