What do CNNs Learn in the First Layer and Why? A Linear Systems Perspective
Rhea Chowers, Yair Weiss
摘要
It has previously been reported that the representation that is learned in the first layer of deep Convolutional Neural Networks (CNNs) is highly consistent across initializations and architectures. In this work, we quantify this consistency by considering the first layer as a filter bank and measuring its energy distribution. We find that the energy distribution is very different from that of the initial weights and is remarkably consistent across random initializations, datasets, architectures and even when the CNNs are trained with random labels. In order to explain this consistency, we derive an analytical formula for the energy profile of linear CNNs and show that this profile is mostly dictated by the second order statistics of image patches in the training set and it will approach a whitening transformation when the number of iterations goes to infinity. Finally, we show that this formula for linear CNNs also gives an excellent fit for the energy profiles learned by commonly used nonlinear CNNs such as ResNet and VGG, and that the first layer of these CNNs indeed perform approximate whitening of their inputs.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper8
- Do Wide and Deep Networks Learn the Same Things? Uncovering How Neural Network Representations Vary with Width and DepthThao Nguyen, Maithra Raghu, Simon KornblithICLR 2021 · 被引用 323 次
- Revisiting Model Stitching to Compare Neural RepresentationsYamini Bansal, Preetum Nakkiran, Boaz BarakNeurIPS 2021 · 被引用 253 次
- Switchable Whitening for Deep Representation LearningXingang Pan, Xiaohang Zhan, Jianping Shi, Xiaoou Tang 等ICCV 2019 · 被引用 204 次
- Similarity and Matching of Neural Network RepresentationsAdrián Csiszárik, Péter Korösi-Szabó, Ákos K. Matszangosz, Gergely Papp 等NeurIPS 2021 · 被引用 105 次
- What Do Neural Networks Learn When Trained With Random Labels?Hartmut Maennel, Ibrahim M. Alabdulmohsin, Ilya O. Tolstikhin, Robert J. N. Baldock 等NeurIPS 2020 · 被引用 99 次
相关 Paper
- A Theory of Contrastive Learning with Natural ImagesAntonio Torralba, Yair WeissICML 2026
- Neural Networks Classify through the Class-Wise Means of Their RepresentationsMohamed El Amine Seddik, Mohamed TamaazoustiAAAI 2022 · 被引用 4 次
- Hierarchical nucleation in deep neural networksDiego Doimo, Aldo Glielmo, Alessio Ansuini, Alessandro LaioNeurIPS 2020 · 被引用 38 次
- Random Matrix Theory Proves that Deep Learning Representations of GAN-data Behave as Gaussian MixturesMohamed El Amine Seddik, Cosme Louart, Mohamed Tamaazousti, Romain CouilletICML 2020 · 被引用 78 次
- Efficient Learning of CNNs using Patch Based FeaturesAlon Brutzkus, Amir Globerson, Eran Malach, Alon Regev Netser 等ICML 2022 · 被引用 6 次
