Lune

AAAI2020Top-tier venue

The HSIC Bottleneck: Deep Learning without Back-Propagation

Kurt Wan-Duo Ma, J. P. Lewis, W. Bastiaan Kleijn

2020Year
180Citations
37Top-tier citations

Abstract

We introduce the HSIC (Hilbert-Schmidt independence criterion) bottleneck for training deep neural networks. The HSIC bottleneck is an alternative to the conventional cross-entropy loss and backpropagation that has a number of distinct advantages. It mitigates exploding and vanishing gradients, resulting in the ability to learn very deep networks without skip connections. There is no requirement for symmetric feedback or update locking. We find that the HSIC bottleneck provides performance on MNIST/FashionMNIST/CIFAR10 classification comparable to backpropagation with a cross-entropy target, even when the system is not encouraged to make the output resemble the classification labels. Appending a single layer trained with SGD (without backpropagation) to reformat the information further improves performance.

Ask about this paper

Your agent reads all of it.

Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.

Questions to start from

Your agent calls

Luneget_paper_fulltext

Ask in Lune

Free to start. No credit card required.

lune papers fulltext 9eb78e8a-7eb0-4711-b2e2-06e0d9e089c4

Cited by top-tier papers37

Ask how each one uses it

Related papers

Dusk over the sea between two cliffs drawn in fine vertical lines