Data ID Extraction Networks for Unsupervised Class- and Classifier-Free Detection of Adversarial Examples

Summary

This study introduces a novel unsupervised method to detect adversarial examples, which are manipulated inputs that fool deep neural networks (DNNs). The approach effectively identifies malicious samples without needing prior knowledge of attack types or data classes.