Deep learning is adaptive to intrinsic dimensionality of model smoothness in anisotropic Besov space
Taiji Suzuki, Atsushi Nitanda
Abstract
Deep learning has exhibited superior performance for various tasks, especially for high-dimensional datasets, such as images. To understand this property, we investigate the approximation and estimation ability of deep learning on anisotropic Besov spaces. The anisotropic Besov space is characterized by direction-dependent smoothness and includes several function classes that have been investigated thus far. We demonstrate that the approximation error and estimation error of deep learning only depend on the average value of the smoothness parameters in all directions. Consequently, the curse of dimensionality can be avoided if the smoothness of the target function is highly anisotropic. Unlike existing studies, our analysis does not require a low-dimensional structure of the input data. We also investigate the minimax optimality of deep learning and compare its performance with that of the kernel method (more generally, linear estimators). The results show that deep learning has better dependence on the input dimensionality if the target function possesses anisotropic smoothness, and it achieves an adaptive rate for functions with spatially inhomogeneous smoothness.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d6a9f833-496a-46b5-a062-8f1bf78b9620Cited by top-tier papers24
- Diffusion Models are Minimax Optimal Distribution EstimatorsKazusato Oko, Shunta Akiyama, Taiji SuzukiICML 2023 · 152 citations
- Black-Box Optimization with Local Generative SurrogatesSergey Shirobokov, Vladislav Belavin, Michael Kagan, Andrey Ustyuzhanin et al.NeurIPS 2020 · 60 citations
- Machine Learning For Elliptic PDEs: Fast Rate Generalization Bound, Neural Scaling Law and Minimax OptimalityYiping Lu, Haoxuan Chen, Jianfeng Lu, Lexing Ying et al.ICLR 2022 · 54 citations
- Besov Function Approximation and Binary Classification on Low-Dimensional Manifolds Using Convolutional Residual NetworksHao Liu, Minshuo Chen, Tuo Zhao, Wenjing LiaoICML 2021 · 42 citations
- Transformers are Minimax Optimal Nonparametric In-Context LearnersJuno Kim, Tai Nakamaki, Taiji SuzukiNeurIPS 2024 · 42 citations
Builds on1
Related papers
- Learnability of convolutional neural networks for infinite dimensional input via mixed and anisotropic smoothnessSho Okumoto, Taiji SuzukiICLR 2022 · 9 citations
- Posterior Contraction for Sparse Neural Networks in Besov Spaces with Intrinsic DimensionalityKyeongwon Lee, Lizhen Lin, Jaewoo Park, Seonghyun JeongNeurIPS 2025 · 4 citations
- Deep Learning meets Nonparametric Regression: Are Weight-Decayed DNNs Locally Adaptive?Kaiqi Zhang, Yu-Xiang WangICLR 2023 · 3 citations
- Minimax optimality of convolutional neural networks for infinite dimensional input-output problems and separation from kernel methodsYuto Nishimura, Taiji SuzukiICLR 2024 · 3 citations
- Approximation with CNNs in Sobolev Space: with Applications to ClassificationGuohao Shen, Yuling Jiao, Yuanyuan Lin, Jian HuangNeurIPS 2022 · 25 citations
