Przeskocz do nawigacji głównej Przeskocz do wyszukiwania Przeskocz do głównej treści

Adversarial Attacks Detection Method for Tabular Data

  • University of Silesia in Katowice
  • Institute of Innovative Technologies EMAG
  • University of Warsaw

Wyniki badań: Wkład do czasopismaArtykułrecenzja

1 Cytowanie z bazy Scopus

Abstrakt

Adversarial attacks involve malicious actors introducing intentional perturbations to machine learning (ML) models, causing unintended behavior. This poses a significant threat to the integrity and trustworthiness of ML models, necessitating the development of robust detection techniques to protect systems from potential threats. The paper proposes a new approach for detecting adversarial attacks using a surrogate model and diagnostic attributes. The method was tested on 22 tabular datasets on which four different ML models were trained. Furthermore, various attacks were conducted, which led to obtaining perturbed data. The proposed approach is characterized by high efficiency in detecting known and unknown attacks—balanced accuracy was above 0.94, with very low false negative rates (0.02–0.10) for binary detection. Sensitivity analysis shows that classifiers trained based on diagnostic attributes can detect even very subtle adversarial attacks.

Język oryginałuangielski
Numer artykułu112
CzasopismoMachine Learning and Knowledge Extraction
Tom7
Numer wydania4
Identyfikatory DOI
Status publikacjiOpublikowano - gru 2025

Obszary tematyczne ASJC Scopus

  • Inżynieria (różne)
  • Sztuczna inteligencja

Fingerprint

Zanurz się w tematy badawcze publikacji „Adversarial Attacks Detection Method for Tabular Data”. Razem tworzą niepowtarzalny odcisk palca.

Cytuj to