Test-Time Detection of Backdoor Triggers for Poisoned Deep Neural   Networks

Xi Li; Zhen Xiang; David J. Miller; George Kesidis

arXiv:2112.03350·cs.CR·December 8, 2021·1 cites

Test-Time Detection of Backdoor Triggers for Poisoned Deep Neural Networks

Xi Li, Zhen Xiang, David J. Miller, George Kesidis

PDF

Open Access

TL;DR

This paper introduces a novel test-time defense mechanism for detecting and identifying backdoor triggers in deep neural networks during inference, addressing a critical gap in existing defenses.

Contribution

It proposes an 'in-flight' detection method that identifies backdoor triggers and infers their source class during testing, unlike prior methods that only detect attacks post-training.

Findings

01

Effective detection of backdoor triggers at test-time

02

Ability to infer source class of detected triggers

03

Demonstrated robustness against various strong backdoor attacks

Abstract

Backdoor (Trojan) attacks are emerging threats against deep neural networks (DNN). A DNN being attacked will predict to an attacker-desired target class whenever a test sample from any source class is embedded with a backdoor pattern; while correctly classifying clean (attack-free) test samples. Existing backdoor defenses have shown success in detecting whether a DNN is attacked and in reverse-engineering the backdoor pattern in a "post-training" regime: the defender has access to the DNN to be inspected and a small, clean dataset collected independently, but has no access to the (possibly poisoned) training set of the DNN. However, these defenses neither catch culprits in the act of triggering the backdoor mapping, nor mitigate the backdoor attack at test-time. In this paper, we propose an "in-flight" defense against backdoor attacks on image classification that 1) detects use of a…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsAdversarial Robustness in Machine Learning · Integrated Circuits and Semiconductor Failure Analysis · Anomaly Detection Techniques and Applications