LISArD: Learning Image Similarity to Defend Against Gray-box Adversarial   Attacks

Joana C. Costa; Tiago Roxo; Hugo Proen\c{c}a; Pedro R. M.; In\'acio

arXiv:2502.20562·cs.CV·March 3, 2025

LISArD: Learning Image Similarity to Defend Against Gray-box Adversarial Attacks

Joana C. Costa, Tiago Roxo, Hugo Proen\c{c}a, Pedro R. M., In\'acio

PDF

Open Access 1 Repo

TL;DR

LISArD is a novel image similarity learning method that enhances robustness against gray-box and white-box adversarial attacks without increasing computational costs, outperforming existing defenses.

Contribution

Proposes LISArD, a new defense mechanism that uses embedding similarity to defend against gray-box attacks, without relying on adversarial training.

Findings

01

LISArD effectively defends against gray-box and white-box attacks.

02

It maintains robustness across multiple neural network architectures.

03

State-of-the-art adversarial distillation models perform poorly without adversarial training.

Abstract

State-of-the-art defense mechanisms are typically evaluated in the context of white-box attacks, which is not realistic, as it assumes the attacker can access the gradients of the target network. To protect against this scenario, Adversarial Training (AT) and Adversarial Distillation (AD) include adversarial examples during the training phase, and Adversarial Purification uses a generative model to reconstruct all the images given to the classifier. This paper considers an even more realistic evaluation scenario: gray-box attacks, which assume that the attacker knows the architecture and the dataset used to train the target network, but cannot access its gradients. We provide empirical evidence that models are vulnerable to gray-box attacks and propose LISArD, a defense mechanism that does not increase computational and temporal costs but provides robustness against gray-box and…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

joana-cabral/lisard
pytorchOfficial

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsAdversarial Robustness in Machine Learning · Generative Adversarial Networks and Image Synthesis · Ethics and Social Impacts of AI