arrow
Return

Adversarial Attacks Detection Method for Tabular Data

delete2025-10-01
delete0
delete
OA
AI
Ł
Łukasz Wawrowski
P
Piotr Biczyk
D
Dominik Ślȩzak
DOI:10.3390/make7040112delete
deleteOriginal
deleteShare
deleteSave
View PDF
Abstract

Abstract

En 中文
Adversarial attacks involve malicious actors introducing intentional perturbations to machine learning (ML) models, causing unintended behavior. This poses a significant threat to the integrity and trustworthiness of ML models, necessitating the development of robust detection techniques to protect systems from potential threats. The paper proposes a new approach for detecting adversarial attacks using a surrogate model and diagnostic attributes. The method was tested on 22 tabular datasets on which four different ML models were trained. Furthermore, various attacks were conducted, which led to obtaining perturbed data. The proposed approach is characterized by high efficiency in detecting known and unknown attacks—balanced accuracy was above 0.94, with very low false negative rates (0.02–0.10) for binary detection. Sensitivity analysis shows that classifiers trained based on diagnostic attributes can detect even very subtle adversarial attacks.
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

M
Machine Learning and Knowledge Extraction
IF:
6
Papers:
795
Citations:
1.8K

Organization

U
university of silesia
Scholars:
362
Papers: 194
Citations: 0
U
University of Warsaw
Scholars:
1.2W
Papers: 1.1W
Citations: 1.1W
S
Silesian University of Technology
Scholars:
6.2K
Papers: 6.2K
Citations: 5.9K
researcher View more organizations