Lune

ICLR2020Top-tier venue

Adversarial Training and Provable Defenses: Bridging the Gap

Mislav Balunovic, Martin T. Vechev

2020Year
186Citations
67Top-tier citations

Abstract

We present COLT, a new method to train neural networks based on a novel combination of adversarial training and provable defenses.The key idea is to model neural network training as a procedure which includes both, the verifier and the adversary.In every iteration, the verifier aims to certify the network using convex relaxation while the adversary tries to find inputs inside that convex relaxation which cause verification to fail.We experimentally show that this training method, named convex layerwise adversarial training (COLT), is promising and achieves the best of both worlds -it produces a state-of-the-art neural network with certified robustness of 60.5% and accuracy of 78.4% on the challenging CIFAR-10 dataset with a 2/255 L perturbation.This significantly improves over the best concurrent results of 54.0% certified robustness and 71.5% accuracy.

Ask about this paper

Ask your agent about it.

Lune has read the top-tier papers around this one, so every answer names the papers it rests on.

Questions to start from

Your agent calls

Lunesearch_papers

Ask in Lune

Free to start. No credit card required.

lune papers get 34b01be4-5329-4220-99f3-9196d548da52

Cited by top-tier papers67

Ask how each one uses it

Related papers

Dusk over the sea between two cliffs drawn in fine vertical lines