The HSIC Bottleneck: Deep Learning without Back-Propagation
Kurt Wan-Duo Ma, J. P. Lewis, W. Bastiaan Kleijn
摘要
We introduce the HSIC (Hilbert-Schmidt independence criterion) bottleneck for training deep neural networks. The HSIC bottleneck is an alternative to the conventional cross-entropy loss and backpropagation that has a number of distinct advantages. It mitigates exploding and vanishing gradients, resulting in the ability to learn very deep networks without skip connections. There is no requirement for symmetric feedback or update locking. We find that the HSIC bottleneck provides performance on MNIST/FashionMNIST/CIFAR10 classification comparable to backpropagation with a cross-entropy target, even when the system is not encouraged to make the output resemble the classification labels. Appending a single layer trained with SGD (without backpropagation) to reformat the information further improves performance.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper37
- Self-Supervised Learning with Kernel Dependence MaximizationYazhe Li, Roman Pogodin, Danica J. Sutherland, Arthur GrettonNeurIPS 2021 · 被引用 107 次
- FwdLLM: Efficient Federated Finetuning of Large Language Models with Perturbed InferencesMengwei Xu, Dongqi Cai, Yaozong Wu, Xiang Li 等USENIX ATC 2024 · 被引用 78 次
- Independent Prototype Propagation for Zero-Shot CompositionalityFrank Ruis, Gertjan J. Burghouts, Doina BucurNeurIPS 2021 · 被引用 77 次
- Towards Crowdsourced Training of Large Neural Networks using Decentralized Mixture-of-ExpertsMax Ryabinin, Anton GusevNeurIPS 2020 · 被引用 71 次
- Demystifying Reasoning Dynamics with Mutual Information: Thinking Tokens are Information Peaks in LLM ReasoningChen Qian, Dongrui Liu, Haochen Wen, Zhen Bai 等NeurIPS 2025 · 被引用 63 次
相关 Paper
- Revisiting Hilbert-Schmidt Information Bottleneck for Adversarial RobustnessZifeng Wang, Tong Jian, Aria Masoomi, Stratis Ioannidis 等NeurIPS 2021 · 被引用 43 次
- Depth-Progressive Monotonic Learning without Global BackpropagationChenhao Ye, Rongguang Ye, Yuchao Zhang, Ming TangICML 2026 · 被引用 1 次
- Kernelized information bottleneck leads to biologically plausible 3-factor Hebbian learning in deep networksRoman Pogodin, Peter E. LathamNeurIPS 2020 · 被引用 48 次
- Towards Continual Learning Desiderata via HSIC-Bottleneck Orthogonalization and Equiangular EmbeddingDepeng Li, Tianqi Wang, Junwei Chen, Qining Ren 等AAAI 2024 · 被引用 10 次
- Deep Isometric Learning for Visual RecognitionHaozhi Qi, Chong You, Xiaolong Wang, Yi Ma 等ICML 2020 · 被引用 57 次
