On the Robustness of Adversarial Training Against Uncertainty Attacks

Emanuele Ledda; Giovanni Scodeller; Daniele Angioni; Giorgio Piras; Antonio Emanuele Cin\`a; Giorgio Fumera; Battista Biggio; Fabio Roli

arXiv:2410.21952·cs.LG·May 28, 2025

On the Robustness of Adversarial Training Against Uncertainty Attacks

Emanuele Ledda, Giovanni Scodeller, Daniele Angioni, Giorgio Piras, Antonio Emanuele Cin\`a, Giorgio Fumera, Battista Biggio, Fabio Roli

PDF

Open Access 1 Repo

TL;DR

This paper investigates how adversarial training enhances the robustness of uncertainty estimates in machine learning models, ensuring more trustworthy outputs under attack scenarios, supported by empirical and theoretical analysis on CIFAR-10 and ImageNet.

Contribution

It demonstrates that defending against adversarial examples inherently improves the security and reliability of uncertainty measures without additional defenses.

Findings

01

Adversarial training leads to more trustworthy uncertainty estimates.

02

Robust models maintain better uncertainty calibration under attack.

03

Empirical validation on CIFAR-10 and ImageNet supports the theoretical claims.

Abstract

In learning problems, the noise inherent to the task at hand hinders the possibility to infer without a certain degree of uncertainty. Quantifying this uncertainty, regardless of its wide use, assumes high relevance for security-sensitive applications. Within these scenarios, it becomes fundamental to guarantee good (i.e., trustworthy) uncertainty measures, which downstream modules can securely employ to drive the final decision-making process. However, an attacker may be interested in forcing the system to produce either (i) highly uncertain outputs jeopardizing the system's availability or (ii) low uncertainty estimates, making the system accept uncertain samples that would instead require a careful inspection (e.g., human intervention). Therefore, it becomes fundamental to understand how to obtain robust uncertainty estimates against these kinds of attacks. In this work, we reveal…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

pralab/UncertaintyAdversarialRobustness
pytorchOfficial

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsAdversarial Robustness in Machine Learning · Fault Detection and Control Systems · Smart Grid Security and Resilience