On the robustness of adversarial training against uncertainty attacks

Ledda, Emanuele
First
;
Angioni, Daniele;Piras, Giorgio;Cinà, Antonio Emanuele;Fumera, Giorgio;Biggio, Battista;Roli, Fabio
Last
2026-01-01

Abstract

In learning problems, the noise inherent to the task at hand hinders the possibility to infer without a certain degree of uncertainty. Quantifying this uncertainty, regardless of its wide use, assumes high relevance for security-sensitive applications. Within these scenarios, it becomes fundamental to guarantee good (i.e., trustworthy) uncertainty measures, which downstream modules can securely employ to drive the final decision-making process. However, an attacker may be interested in forcing the system to produce either (i) highly uncertain outputs jeopardizing the system’s availability or (ii) low uncertainty estimates, making the system accept uncertain samples that would instead require a careful inspection (e.g., human intervention). Therefore, it becomes fundamental to understand how to obtain robust uncertainty estimates against these kinds of attacks. In this work, we reveal both empirically and theoretically that defending against adversarial examples, i.e., carefully perturbed samples that cause misclassification, additionally guarantees a more secure, trustworthy uncertainty estimate under common attack scenarios without the need for an ad-hoc defense strategy. To support our claims, we evaluate multiple adversarial-robust classification models from the publicly available benchmark RobustBench on the CIFAR-10 and ImageNet datasets, and on a robust semantic segmentation model evaluated on Pascal-VOC. The code for the reproducibility of the experiments is available at the following link:https://github.com/pralab/UncertaintyAdversarialRobustness.
2026
2025
Inglese
172
112519
1
13
13
Esperti anonimi
internazionale
scientifica
Uncertainty quantification; Adversarial machine learning; Neural networks
no
Ledda, Emanuele; Scodeller, Giovanni; Angioni, Daniele; Piras, Giorgio; Cinà, Antonio Emanuele; Fumera, Giorgio; Biggio, Battista; Roli, Fabio ...espandi
1.1 Articolo in rivista
info:eu-repo/semantics/article
1 Contributo su Rivista::1.1 Articolo in rivista
262
8
partially_open
   European Lighthouse on Secure and Safe AI
   ELSA
   European Commission
   Horizon Europe Framework Programme
   101070617

   A COMPREHENSIVE TRUSTWORTHY FRAMEWORK FOR CONNECTED MACHINE LEARNING AND SECURE INTERCONNECTED AI SOLUTIONS
   CoEvolution
   European Commission
   Horizon Europe Framework Programme
   101168560
Files in This Item:
File Size Format  
1-s2.0-S0031320325011823-main.pdf

open access

Type: versione editoriale
Size 7.16 MB
Format Adobe PDF
7.16 MB Adobe PDF View/Open
2410.21952v2.pdf

Solo gestori archivio

Type: versione pre-print
Size 973.15 kB
Format Adobe PDF
973.15 kB Adobe PDF & nbsp; View / Open   Request a copy

Items in DSpace are protected by copyright, with all rights reserved, unless otherwise indicated.

Questionnaire and social

Share on:
Impostazioni cookie