Implicit Bias of Linear Equivariant Networks
Hannah Lawrence, Bobak Toussi Kiani, Kristian G. Georgiev, Andrew K. Dienes
摘要
Group equivariant convolutional neural networks (G-CNNs) are generalizations of convolutional neural networks (CNNs) which excel in a wide range of technical applications by explicitly encoding symmetries, such as rotations and permutations, in their architectures. Although the success of G-CNNs is driven by their explicit symmetry bias, a recent line of work has proposed that the implicit bias of training algorithms on particular architectures is key to understanding generalization for overparameterized neural nets. In this context, we show that -layer full-width linear G-CNNs trained via gradient descent for binary classification converge to solutions with low-rank Fourier matrix coefficients, regularized by the -Schatten matrix norm. Our work strictly generalizes previous analysis on the implicit bias of linear CNNs to linear G-CNNs over all finite groups, including the challenging setting of non-commutative groups (such as permutations), as well as band-limited G-CNNs over infinite groups. We validate our theorems via experiments on a variety of groups, and empirically explore more realistic nonlinear networks, which locally capture similar regularization patterns. Finally, we provide intuitive interpretations of our Fourier space implicit regularization results in real space via uncertainty principles.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- projUNN: efficient method for training deep networks with unitary matricesBobak Toussi Kiani, Randall Balestriero, Yann LeCun, Seth LloydNeurIPS 2022 · 被引用 42 次
- The Exact Sample Complexity Gain from Invariances for Kernel RegressionBehrooz Tahmasebi, Stefanie JegelkaNeurIPS 2023 · 被引用 29 次
- On the hardness of learning under symmetriesBobak T. Kiani, Thien Le, Hannah Lawrence, Stefanie Jegelka 等ICLR 2024 · 被引用 14 次
- A Universal Class of Sharpness-Aware Minimization AlgorithmsBehrooz Tahmasebi, Ashkan Soleymani, Dara Bahri, Stefanie Jegelka 等ICML 2024 · 被引用 13 次
- Spectral Graph Neural Networks are Incomplete on Graphs with a Simple SpectrumSnir Hordan, Maya Bechler-Speicher, Gur Lifshitz, Nadav DymNeurIPS 2025 · 被引用 5 次
它引用的顶会 Paper4
- Gradient Descent Maximizes the Margin of Homogeneous Neural NetworksKaifeng Lyu, Jian LiICLR 2020 · 被引用 402 次
- Implicit Regularization in Deep Learning May Not Be Explainable by NormsNoam Razin, Nadav CohenNeurIPS 2020 · 被引用 178 次
- Provably Strict Generalisation Benefit for Equivariant ModelsBryn Elesedy, Sheheryar ZaidiICML 2021 · 被引用 100 次
- A unifying view on implicit bias in training linear neural networksChulhee Yun, Shankar Krishnan, Hossein MobahiICLR 2021 · 被引用 94 次
相关 Paper
- Group Downsampling with Equivariant Anti-aliasingMd Ashiqur Rahman, Raymond A. YehICLR 2025
- Universal Equivariant Multilayer PerceptronsSiamak RavanbakhshICML 2020 · 被引用 60 次
- Learning Partial Equivariances From DataDavid W. Romero, Suhas LohitNeurIPS 2022 · 被引用 54 次
- Learning symmetries via weight-sharing with doubly stochastic tensorsPutri A. van der Linden, Alejandro García-Castellanos, Sharvaree P. Vadgama, Thijs P. Kuipers 等NeurIPS 2024 · 被引用 6 次
- B-Spline CNNs on Lie groupsErik J. BekkersICLR 2020 · 被引用 155 次
