Role of Locality and Weight Sharing in Image-Based Tasks: A Sample Complexity Separation between CNNs, LCNs, and FCNs
Aakash Lahoti, Stefani Karp, Ezra Winston, Aarti Singh, Yuanzhi Li
摘要
Vision tasks are characterized by the properties of locality and translation invariance. The superior performance of convolutional neural networks (CNNs) on these tasks is widely attributed to the inductive bias of locality and weight sharing baked into their architecture. Existing attempts to quantify the statistical benefits of these biases in CNNs over locally connected convolutional neural networks (LCNs) and fully connected neural networks (FCNs) fall into one of the following categories: either they disregard the optimizer and only provide uniform convergence upper bounds with no separating lower bounds, or they consider simplistic tasks that do not truly mirror the locality and translation invariance as found in real-world vision tasks. To address these deficiencies, we introduce the Dynamic Signal Distribution (DSD) classification task that models an image as consisting of patches, each of dimension , and the label is determined by a -sparse signal vector that can freely appear in any one of the patches. On this task, for any orthogonally equivariant algorithm like gradient descent, we prove that CNNs require samples, whereas LCNs require samples, establishing the statistical advantages of weight sharing in translation invariant tasks. Furthermore, LCNs need samples, compared to samples for FCNs, showcasing the benefits of locality in local tasks. Additionally, we develop information theoretic tools for analyzing randomized algorithms, which may be of interest for statistical research.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper6
- ViM: Out-Of-Distribution with Virtual-logit MatchingHaoqi Wang, Zhizhong Li, Litong Feng, Wayne ZhangCVPR 2022 · 被引用 227 次
- Generalization bounds for deep convolutional neural networksPhilip M. Long, Hanie SedghiICLR 2020 · 被引用 102 次
- Local Signal Adaptivity: Provable Feature Learning in Neural Networks Beyond KernelsStefani Karp, Ezra Winston, Yuanzhi Li, Aarti SinghNeurIPS 2021 · 被引用 38 次
- Computational Separation Between Convolutional and Fully-Connected NetworksEran Malach, Shai Shalev-ShwartzICLR 2021 · 被引用 32 次
- Why Are Convolutional Nets More Sample-Efficient than Fully-Connected Nets?Zhiyuan Li, Yi Zhang, Sanjeev AroraICLR 2021 · 被引用 22 次
相关 Paper
- Theoretical Analysis of the Inductive Biases in Deep Convolutional NetworksZihao Wang, Lei WuNeurIPS 2023 · 被引用 10 次
- On the Connection between Local Attention and Dynamic Depth-wise ConvolutionQi Han, Zejia Fan, Qi Dai, Lei Sun 等ICLR 2022 · 被引用 144 次
- Revisiting Spatial Invariance with Low-Rank Local ConnectivityGamaleldin F. Elsayed, Prajit Ramachandran, Jonathon Shlens, Simon KornblithICML 2020 · 被引用 51 次
- Towards Biologically Plausible Convolutional NetworksRoman Pogodin, Yash Mehta, Timothy P. Lillicrap, Peter E. LathamNeurIPS 2021 · 被引用 30 次
- A Fourier perspective on the learning dynamics of neural networks: from sample complexities to mechanistic insightsFabiola Ricci, Claudia Merger, Sebastian GoldtICML 2026
