Adversarial Training and Robustness for Multiple Perturbations

Florian Tram\`er; Dan Boneh

arXiv:1904.13000·cs.LG·October 21, 2019·84 cites

Adversarial Training and Robustness for Multiple Perturbations

Florian Tram\`er, Dan Boneh

PDF

Open Access 1 Repo

TL;DR

This paper investigates the inherent trade-offs in training models to be robust against multiple types of adversarial perturbations, revealing fundamental limitations and challenges in achieving comprehensive robustness.

Contribution

The study provides formal proof of robustness trade-offs, introduces new multi-perturbation training schemes, and demonstrates the limitations of current adversarial training methods across multiple perturbation types.

Findings

01

Robustness trade-offs exist between different perturbation types.

02

Models trained on multiple attacks perform worse than those trained on individual attacks.

03

Gradient masking significantly reduces adversarial accuracy on MNIST.

Abstract

Defenses against adversarial examples, such as adversarial training, are typically tailored to a single perturbation type (e.g., small $ℓ_{\infty}$ -noise). For other perturbations, these defenses offer no guarantees and, at times, even increase the model's vulnerability. Our aim is to understand the reasons underlying this robustness trade-off, and to train models that are simultaneously robust to multiple perturbation types. We prove that a trade-off in robustness to different types of $ℓ_{p}$ -bounded and spatial perturbations must exist in a natural and simple statistical setting. We corroborate our formal analysis by demonstrating similar robustness trade-offs on MNIST and CIFAR10. Building upon new multi-perturbation adversarial training schemes, and a novel efficient attack for finding $ℓ_{1}$ -bounded adversarial examples, we show that no model trained against multiple attacks…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

ftramer/MultiRobustness
tfOfficial

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsAdversarial Robustness in Machine Learning · Bacillus and Francisella bacterial research