arrow
Return

Detecting adversarial manipulation using inductive Venn-ABERS predictors

delete2020-11-01
delete5
delete
OA
AI
J
Jonathan Peck *
B
Bart Goossens
Y
Yvan Saeys
DOI:10.1016/j.neucom.2019.11.113delete
deleteOriginal
deleteShare
deleteSave
View PDF
Abstract

Abstract

En 中文
Inductive Venn-ABERS predictors (IVAPs) are a type of probabilistic predictors with the theoretical guarantee that their predictions are perfectly calibrated. In this paper, we propose to exploit this calibration property for the detection of adversarial examples in binary classification tasks. By rejecting predictions if the uncertainty of the IVAP is too high, we obtain an algorithm that is both accurate on the original test set and resistant to adversarial examples. This robustness is observed on adversarials for the underlying model as well as adversarials that were generated by taking the IVAP into account. The method appears to offer competitive robustness compared to the state-of-the-art in adversarial defense yet it is computationally much more tractable. (C) 2020 The Author(s). Published by Elsevier B.V.
Keywords:
Adversarial robustness
Conformal prediction
Supervised learning
Deep learning
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

Neurocomputing cover
Neurocomputing
IF:
6.5
Papers:
2.5W
Citations:
6.5W

Organization

G
Ghent University
Scholars:
5.2W
Papers: 4.5W
Citations: 5.5W