WLD-Reg: A Data-Dependent Within-Layer Diversity Regularizer
Firas Laakom, Jenni Raitoharju, Alexandros Iosifidis, Moncef Gabbouj
摘要
Neural networks are composed of multiple layers arranged in a hierarchical structure jointly trained with a gradient-based optimization, where the errors are back-propagated from the last layer back to the first one. At each optimization step, neurons at a given layer receive feedback from neurons belonging to higher layers of the hierarchy. In this paper, we propose to complement this traditional 'between-layer' feedback with additional 'within-layer' feedback to encourage the diversity of the activations within the same layer. To this end, we measure the pairwise similarity between the outputs of the neurons and use it to model the layer's overall diversity. We present an extensive empirical study confirming that the proposed approach enhances the performance of several state-of-the-art neural network models in multiple tasks. The code is publically available at https://github.com/firasl/AAAI-23-WLD-Reg.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Project and Probe: Sample-Efficient Adaptation by Interpolating Orthogonal FeaturesAnnie S. Chen, Yoonho Lee, Amrith Setlur, Sergey Levine 等ICLR 2024 · 被引用 5 次
- Test-time Diverse Reasoning by Riemannian Activation SteeringLy Tran Ho Khanh, Dongxuan Zhu, Man-Chung Yue, Viet Anh NguyenAAAI 2026 · 被引用 1 次
它引用的顶会 Paper6
- MLP-Mixer: An all-MLP Architecture for VisionIlya O. Tolstikhin, Neil Houlsby, Alexander Kolesnikov, Lucas Beyer 等NeurIPS 2021 · 被引用 3,862 次
- Barlow Twins: Self-Supervised Learning via Redundancy ReductionJure Zbontar, Li Jing, Ishan Misra, Yann LeCun 等ICML 2021 · 被引用 2,942 次
- Sharpness-aware Minimization for Efficiently Improving GeneralizationPierre Foret, Ariel Kleiner, Hossein Mobahi, Behnam NeyshaburICLR 2021 · 被引用 1,861 次
- Deep Double Descent: Where Bigger Models and More Data HurtPreetum Nakkiran, Gal Kaplun, Yamini Bansal, Tristan Yang 等ICLR 2020 · 被引用 1,108 次
- Pay Attention to MLPsHanxiao Liu, Zihang Dai, David R. So, Quoc V. LeNeurIPS 2021 · 被引用 912 次
相关 Paper
- Lamina-specific neuronal properties promote robust, stable signal propagation in feedforward networksDongqi Han, Erik De Schutter, Sungho HongNeurIPS 2020 · 被引用 3 次
- MMA Regularization: Decorrelating Weights of Neural Networks by Maximizing the Minimal AnglesZhennan Wang, Canqun Xiang, Wenbin Zou, Chen XuNeurIPS 2020 · 被引用 25 次
- Neuron with Steady Response Leads to Better GeneralizationQiang Fu, Lun Du, Haitao Mao, Xu Chen 等NeurIPS 2022 · 被引用 5 次
- Fixed-Weight Difference Target PropagationTatsukichi Shibuya, Nakamasa Inoue, Rei Kawakami, Ikuro SatoAAAI 2023 · 被引用 6 次
- Intraclass clustering: an implicit learning ability that regularizes DNNsSimon Carbonnelle, Christophe De VleeschouwerICLR 2021 · 被引用 2 次
