Spectral Bias in Practice: The Role of Function Frequency in Generalization
Sara Fridovich-Keil, Raphael Gontijo Lopes, Rebecca Roelofs
摘要
Despite their ability to represent highly expressive functions, deep learning models seem to find simple solutions that generalize surprisingly well. Spectral biasthe tendency of neural networks to prioritize learning low frequency functions -is one possible explanation for this phenomenon, but so far spectral bias has primarily been observed in theoretical models and simplified experiments. In this work, we propose methodologies for measuring spectral bias in modern image classification networks on CIFAR-10 and ImageNet. We find that these networks indeed exhibit spectral bias, and that interventions that improve test accuracy on CIFAR-10 tend to produce learned functions that have higher frequencies overall but lower frequencies in the vicinity of examples from each class. This trend holds across variation in training time, model architecture, number of training examples, data augmentation, and self-distillation. We also explore the connections between function frequency and image frequency and find that spectral bias is sensitive to the low frequencies prevalent in natural images. On ImageNet, we find that learned function frequency also varies with internal class diversity, with higher frequencies on more diverse classes. Our work enables measuring and ultimately influencing the spectral behavior of neural networks used for image classification, and is a step towards understanding why deep models generalize well. * Work done as an intern and student researcher at Google Brain.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper15
- Spectral Co-Distillation for Personalized Federated LearningZihan Chen, Howard H. Yang, Tony Q. S. Quek, Kai Fong Ernest ChongNeurIPS 2023 · 被引用 30 次
- Spectral Convolutional Conditional Neural ProcessesPeiman Mohseni, Nick DuffieldNeurIPS 2025 · 被引用 10 次
- Decompose to Understand, Fuse to Detect: Frequency-Decoupled Anomaly Detection for Encrypted Network TrafficXinglin Lian, Chengtai Cao, Ting Zhong, Yong Wang 等INFOCOM 2026 · 被引用 8 次
- Why are Sensitive Functions Hard for Transformers?Michael Hahn, Mark RofinACL 2024 · 被引用 3 次
- Models Out of Line: A Fourier Lens on Distribution Shift RobustnessSara Fridovich-Keil, Brian R. Bartoldson, James Diffenderfer, Bhavya Kailkhura 等NeurIPS 2022 · 被引用 2 次
它引用的顶会 Paper15
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Training data-efficient image transformers & distillation through attentionHugo Touvron, Matthieu Cord, Matthijs Douze, Francisco Massa 等ICML 2021 · 被引用 8,974 次
- RandAugment: Practical Automated Data Augmentation with a Reduced Search SpaceEkin Dogus Cubuk, Barret Zoph, Jonathon Shlens, Quoc LeNeurIPS 2020 · 被引用 4,453 次
- CoAtNet: Marrying Convolution and Attention for All Data SizesZihang Dai, Hanxiao Liu, Quoc V. Le, Mingxing TanNeurIPS 2021 · 被引用 1,747 次
- The Pitfalls of Simplicity Bias in Neural NetworksHarshay Shah, Kaustav Tamuly, Aditi Raghunathan, Prateek Jain 等NeurIPS 2020 · 被引用 503 次
相关 Paper
- A Computable Definition of the Spectral BiasJonas Kiessling, Filip ThorAAAI 2022 · 被引用 14 次
- Addressing Spectral Bias of Deep Neural Networks by Multi-Grade Deep LearningRonglong Fang, Yuesheng XuNeurIPS 2024 · 被引用 22 次
- An Inductive Bias for Tabular Deep LearningEge Beyazit, Jonathan Kozaczuk, Bo Li, Vanessa Wallace 等NeurIPS 2023 · 被引用 26 次
- A Fourier perspective on the learning dynamics of neural networks: from sample complexities to mechanistic insightsFabiola Ricci, Claudia Merger, Sebastian GoldtICML 2026
- Learning in the Frequency DomainKai Xu, Minghai Qin, Fei Sun, Yuhao Wang 等CVPR 2020
