Deep learning is adaptive to intrinsic dimensionality of model smoothness in anisotropic Besov space
Taiji Suzuki, Atsushi Nitanda
摘要
Deep learning has exhibited superior performance for various tasks, especially for high-dimensional datasets, such as images. To understand this property, we investigate the approximation and estimation ability of deep learning on anisotropic Besov spaces. The anisotropic Besov space is characterized by direction-dependent smoothness and includes several function classes that have been investigated thus far. We demonstrate that the approximation error and estimation error of deep learning only depend on the average value of the smoothness parameters in all directions. Consequently, the curse of dimensionality can be avoided if the smoothness of the target function is highly anisotropic. Unlike existing studies, our analysis does not require a low-dimensional structure of the input data. We also investigate the minimax optimality of deep learning and compare its performance with that of the kernel method (more generally, linear estimators). The results show that deep learning has better dependence on the input dimensionality if the target function possesses anisotropic smoothness, and it achieves an adaptive rate for functions with spatially inhomogeneous smoothness.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper24
- Diffusion Models are Minimax Optimal Distribution EstimatorsKazusato Oko, Shunta Akiyama, Taiji SuzukiICML 2023 · 被引用 152 次
- Black-Box Optimization with Local Generative SurrogatesSergey Shirobokov, Vladislav Belavin, Michael Kagan, Andrey Ustyuzhanin 等NeurIPS 2020 · 被引用 60 次
- Machine Learning For Elliptic PDEs: Fast Rate Generalization Bound, Neural Scaling Law and Minimax OptimalityYiping Lu, Haoxuan Chen, Jianfeng Lu, Lexing Ying 等ICLR 2022 · 被引用 54 次
- Besov Function Approximation and Binary Classification on Low-Dimensional Manifolds Using Convolutional Residual NetworksHao Liu, Minshuo Chen, Tuo Zhao, Wenjing LiaoICML 2021 · 被引用 42 次
- Transformers are Minimax Optimal Nonparametric In-Context LearnersJuno Kim, Tai Nakamaki, Taiji SuzukiNeurIPS 2024 · 被引用 42 次
它引用的顶会 Paper1
相关 Paper
- Learnability of convolutional neural networks for infinite dimensional input via mixed and anisotropic smoothnessSho Okumoto, Taiji SuzukiICLR 2022 · 被引用 9 次
- Posterior Contraction for Sparse Neural Networks in Besov Spaces with Intrinsic DimensionalityKyeongwon Lee, Lizhen Lin, Jaewoo Park, Seonghyun JeongNeurIPS 2025 · 被引用 4 次
- Deep Learning meets Nonparametric Regression: Are Weight-Decayed DNNs Locally Adaptive?Kaiqi Zhang, Yu-Xiang WangICLR 2023 · 被引用 3 次
- Minimax optimality of convolutional neural networks for infinite dimensional input-output problems and separation from kernel methodsYuto Nishimura, Taiji SuzukiICLR 2024 · 被引用 3 次
- Approximation with CNNs in Sobolev Space: with Applications to ClassificationGuohao Shen, Yuling Jiao, Yuanyuan Lin, Jian HuangNeurIPS 2022 · 被引用 25 次
