On the Transferability of Adversarial Examples between Encrypted Models

Miki Tanaka; Isao Echizen; Hitoshi Kiya

arXiv:2209.02997·cs.CV·September 8, 2022

On the Transferability of Adversarial Examples between Encrypted Models

Miki Tanaka, Isao Echizen, Hitoshi Kiya

PDF

Open Access

TL;DR

This paper investigates whether encrypting models for adversarial robustness affects the transferability of adversarial examples, finding that encryption not only enhances robustness but also reduces transferability.

Contribution

It is the first study to analyze the transferability of adversarial examples between encrypted models, revealing encryption's impact on transferability and robustness.

Findings

01

Encrypted models are more robust against adversarial examples.

02

Encryption reduces the transferability of adversarial examples.

03

AutoAttack confirms robustness and transferability reduction.

Abstract

Deep neural networks (DNNs) are well known to be vulnerable to adversarial examples (AEs). In addition, AEs have adversarial transferability, namely, AEs generated for a source model fool other (target) models. In this paper, we investigate the transferability of models encrypted for adversarially robust defense for the first time. To objectively verify the property of transferability, the robustness of models is evaluated by using a benchmark attack method, called AutoAttack. In an image-classification experiment, the use of encrypted models is confirmed not only to be robust against AEs but to also reduce the influence of AEs in terms of the transferability of models.

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsAdversarial Robustness in Machine Learning