Adversarial Training and Provable Robustness: A Tale of Two Objectives

Jiameng Fan; Wenchao Li

arXiv:2008.06081·cs.LG·June 8, 2021·1 cites

Adversarial Training and Provable Robustness: A Tale of Two Objectives

Jiameng Fan, Wenchao Li

PDF

Open Access 1 Repo

TL;DR

This paper introduces a unified framework combining adversarial training and provable robustness verification, employing a novel gradient technique to improve neural network robustness with strong empirical results on MNIST and CIFAR-10.

Contribution

It presents a joint optimization approach with a new gradient method that enhances certifiable robustness, outperforming prior methods in provable adversarial defense.

Findings

01

Achieved 6.60% verified test error on MNIST at epsilon=0.3

02

Achieved 66.57% verified test error on CIFAR-10 at epsilon=8/255

03

Method outperforms existing approaches in provable robustness

Abstract

We propose a principled framework that combines adversarial training and provable robustness verification for training certifiably robust neural networks. We formulate the training problem as a joint optimization problem with both empirical and provable robustness objectives and develop a novel gradient-descent technique that can eliminate bias in stochastic multi-gradients. We perform both theoretical analysis on the convergence of the proposed technique and experimental comparison with state-of-the-arts. Results on MNIST and CIFAR-10 show that our method can consistently match or outperform prior approaches for provable l infinity robustness. Notably, we achieve 6.60% verified test error on MNIST at epsilon = 0.3, and 66.57% on CIFAR-10 with epsilon = 8/255.

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

JmfanBU/AdvIBP
pytorchOfficial

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsAdversarial Robustness in Machine Learning · Advanced Neural Network Applications · Explainable Artificial Intelligence (XAI)