Stateful Detection of Adversarial Reprogramming

Yang Zheng; Xiaoyi Feng; Zhaoqiang Xia; Xiaoyue Jiang; Maura Pintor,; Ambra Demontis; Battista Biggio; Fabio Roli

arXiv:2211.02885·cs.GT·November 8, 2022

Stateful Detection of Adversarial Reprogramming

Yang Zheng, Xiaoyi Feng, Zhaoqiang Xia, Xiaoyue Jiang, Maura Pintor,, Ambra Demontis, Battista Biggio, Fabio Roli

PDF

Open Access

TL;DR

This paper introduces a stateful detection method for adversarial reprogramming attacks on machine learning models, effectively identifying malicious queries and reducing attack feasibility even with adaptive strategies.

Contribution

The paper demonstrates for the first time that stateful defenses can detect adversarial reprogramming attacks, including adaptive ones, in black-box scenarios.

Findings

01

Stateful defenses can detect adversarial reprogramming attacks.

02

Detection remains effective even when attackers fine-tune adversarial programs.

03

Blocking malicious users reduces attack success.

Abstract

Adversarial reprogramming allows stealing computational resources by repurposing machine learning models to perform a different task chosen by the attacker. For example, a model trained to recognize images of animals can be reprogrammed to recognize medical images by embedding an adversarial program in the images provided as inputs. This attack can be perpetrated even if the target model is a black box, supposed that the machine-learning model is provided as a service and the attacker can query the model and collect its outputs. So far, no defense has been demonstrated effective in this scenario. We show for the first time that this attack is detectable using stateful defenses, which store the queries made to the classifier and detect the abnormal cases in which they are similar. Once a malicious query is detected, the account of the user who made it can be blocked. Thus, the attacker…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsAdversarial Robustness in Machine Learning · Physical Unclonable Functions (PUFs) and Hardware Security