Linear CNNs Discover the Statistical Structure of the Dataset Using Only the Most Dominant Frequencies
Hannah Pinson, Joeri Lenaerts, Vincent Ginis
Abstract
We here present a stepping stone towards a deeper understanding of convolutional neural networks (CNNs) in the form of a theory of learning in linear CNNs. Through analyzing the gradient descent equations, we discover that the evolution of the network during training is determined by the interplay between the dataset structure and the convolutional network structure. We show that linear CNNs discover the statistical structure of the dataset with non-linear, ordered, stage-like transitions, and that the speed of discovery changes depending on the relationship between the dataset and the convolutional network structure. Moreover, we find that this interplay lies at the heart of what we call the "dominant frequency bias", where linear CNNs arrive at these discoveries using only the dominant frequencies of the different structural parts present in the dataset. We furthermore provide experiments that show how our theory relates to deep, non-linear CNNs used in practice. Our findings shed new light on the inner working of CNNs, and can help explain their shortcut learning and their tendency to rely on texture instead of shape. In addition to a neural network's pre-defined architecture, the parameters of the network obtain an implicit structure during training. For example, it has been shown that weight matrices can exhibit structural patterns, such as clusters and branches (Voss et al., 2021; Casper et al., 2022) . On the other hand, the input dataset also has an implicit structure arising from patterns and relationships between the samples. E.g., in a classification task, dogs are more visually similar
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0308fd3a-0af5-4bcf-a281-918581f7f536Cited by top-tier papers5
- Towards Combating Frequency Simplicity-biased Learning for Domain GeneralizationXilin He, Jingyu Hu, Qinliang Lin, Cheng Luo et al.NeurIPS 2024 · 16 citations
- Convolutions and More as Einsum: A Tensor Network Perspective with Advances for Second-Order MethodsFelix DangelNeurIPS 2024 · 5 citations
- A Solvable Attention for Neural Scaling LawsBochen Lyu, Di Wang, Zhanxing ZhuICLR 2025
- Domain Adaptive Object Detection via Dynamic Causal RefinementZeyu Ma, Jiaqi Huang, Yitong Qin, Ziqiang Zheng et al.ICML 2026
- FreqDebias: Towards Generalizable Deepfake Detection via Consistency-Driven Frequency DebiasingHossein Kashiani, Niloufar Alipour Talemi, Fatemeh AfghahCVPR 2025
Builds on5
- Frequency Bias in Neural Networks for Input of Non-Uniform DensityRonen Basri, Meirav Galun, Amnon Geifman, David W. Jacobs et al.ICML 2020 · 229 citations
- Neural Networks as Kernel Learners: The Silent Alignment EffectAlexander B. Atanasov, Blake Bordelon, Cengiz PehlevanICLR 2022 · 110 citations
- Exact learning dynamics of deep linear networks with prior knowledgeLukas Braun, Clémentine C. J. Dominé, James Fitzgerald, Andrew M. SaxeNeurIPS 2022 · 75 citations
- Neural networks trained with SGD learn distributions of increasing complexityMaria Refinetti, Alessandro Ingrosso, Sebastian GoldtICML 2023 · 58 citations
- The dynamics of representation learning in shallow, non-linear autoencodersMaria Refinetti, Sebastian GoldtICML 2022 · 25 citations
Related papers
- Which Layer is Learning Faster? A Systematic Exploration of Layer-wise Convergence Rate for Deep Neural NetworksYixiong Chen, Alan L. Yuille, Zongwei ZhouICLR 2023
- Shape or Texture: Understanding Discriminative Features in CNNsMd. Amirul Islam, Matthew Kowal, Patrick Esser, Sen Jia et al.ICLR 2021 · 86 citations
- What do neural networks learn in image classification? A frequency shortcut perspectiveShunxin Wang, Raymond N. J. Veldhuis, Christoph Brune, Nicola StrisciuglioICCV 2023 · 51 citations
- Deep Frequency Principle Towards Understanding Why Deeper Learning Is FasterZhiqin John Xu, Hanxu ZhouAAAI 2021 · 67 citations
- Implicit Bias of Linear Equivariant NetworksHannah Lawrence, Bobak Toussi Kiani, Kristian G. Georgiev, Andrew K. DienesICML 2022 · 18 citations
