Related Experiment Video
Updated: May 4, 2026

Evidence-based Knowledge Synthesis and Hypothesis Validation: Navigating Biomedical Knowledge Bases via Explainable AI and Agentic Systems
Published on: June 13, 2025
Hybrid reverse knowledge distillation for adversarial example detection
Hyun Kwon1, Joo Bon Maeng1, Dae-Jin Kim2
1Department of Artificial Intelligence and Data Science, Korea Military Academy, Seoul, 01805, South Korea.
None:
Deep neural networks have achieved remarkable success in various computer vision tasks, yet they remain vulnerable to adversarial examples-carefully crafted perturbations that are imperceptible to humans but cause misclassification. Detecting such adversarial inputs is crucial for deploying reliable AI systems, particularly in safety-critical applications such as medical diagnosis and autonomous driving. In this paper, we propose a novel adversarial detection method based on Hybrid Reverse Knowledge Distillation (Hybrid RKD). Our approach trains a multi-scale decoder to reconstruct intermediate feature representations of a frozen teacher encoder using only clean images. The key insight is that adversarial perturbations cause feature-level distortions that the decoder, trained exclusively on normal data, cannot accurately reconstruct. We further enhance detection performance by incorporating Mahalanobis-style statistical distance metrics that capture distribution-level anomalies. Extensive experiments on CIFAR-10 and ISIC2018 datasets demonstrate that our method achieves state-of-the-art detection performance against various adversarial attacks including FGSM, PGD, DeepFool, C&W, and AutoAttack, with average AUROC scores of 0.790 and 0.929 across five attack categories, respectively. Our hybrid approach consistently outperforms existing detection methods including LID, Mahalanobis, and ODIN, while requiring no adversarial examples during training.
Related Concept Videos
Deductive Reasoning
For example, a researcher can deduce specific predictions...
Hindsight Biases
Difference from Background: Limit of Detection
The LOD indicates the presence or absence...
Associative Learning
Classical conditioning, also known...
Generalization, Discrimination, and Extinction
Generalization occurs when a behavior reinforced in one context is performed in similar situations. For instance, a student who studies diligently for calculus and receives excellent grades might apply the same study habits to psychology and history, expecting similar results. Generalization shows how learning in one setting can influence behavior in...
Understanding Deception