Defense Against Adversarial Attacks using Convolutional Auto-Encoders

Shreyasi Mandal

arXiv:2312.03520·cs.CV·December 7, 2023·2 cites

Defense Against Adversarial Attacks using Convolutional Auto-Encoders

Shreyasi Mandal

PDF

Open Access 1 Repo

TL;DR

This paper proposes a convolutional autoencoder-based method to improve the robustness of deep learning classifiers against adversarial attacks by restoring perturbed inputs.

Contribution

It introduces a novel autoencoder approach specifically designed to counteract adversarial perturbations in input data.

Findings

01

Enhanced model robustness against adversarial attacks

02

Restoration of input images improves classification accuracy

03

Effective suppression of imperceptible perturbations

Abstract

Deep learning models, while achieving state-of-the-art performance on many tasks, are susceptible to adversarial attacks that exploit inherent vulnerabilities in their architectures. Adversarial attacks manipulate the input data with imperceptible perturbations, causing the model to misclassify the data or produce erroneous outputs. This work is based on enhancing the robustness of targeted classifier models against adversarial attacks. To achieve this, an convolutional autoencoder-based approach is employed that effectively counters adversarial perturbations introduced to the input images. By generating images closely resembling the input images, the proposed methodology aims to restore the model's accuracy.

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

Shreyasi2002/Adversarial_Attack_Defense
pytorchOfficial

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsAdversarial Robustness in Machine Learning

MethodsSolana Customer Service Number +1-833-534-1729