Lune

ICML2020Top-tier venue

Efficient Proximal Mapping of the 1-path-norm of Shallow Networks

Fabian Latorre, Paul Rolland, Nadav Hallak, Volkan Cevher

2020Year
4Citations

Abstract

We demonstrate two new important properties of the 1-path-norm of shallow neural networks. First, despite its non-smoothness and non-convexity it allows a closed form proximal operator which can be efficiently computed, allowing the use of stochastic proximal-gradient-type methods for regularized empirical risk minimization. Second, when the activation functions is differentiable, it provides an upper bound on the Lipschitz constant of the network. Such bound is tighter than the trivial layer-wise product of Lipschitz constants, motivating its use for training networks robust to adversarial perturbations. In practical experiments we illustrate the advantages of using the proximal mapping and we compare the robustness-accuracy trade-off induced by the 1path-norm, L1-norm and layer-wise constraints on the Lipschitz constant (Parseval networks). * Equal contribution 1 Learning, information and optimization systems laboratory (LIONS), EPFL, Switzerland.

Ask about this paper

Your agent reads all of it.

Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.

Questions to start from

Your agent calls

Luneget_paper_fulltext

Ask in Lune

Free to start. No credit card required.

lune papers fulltext b4378b63-ad9f-4eb8-88e0-1802afb8d0c0

Builds on2

Related papers

Dusk over the sea between two cliffs drawn in fine vertical lines