Preprint
Jul 2026
Adversarial Deepfake Generation and an Investigation of Purification-Based Adversarial Detection
A self-initiated investigation of purification-based adversarial detection, comparing three families of detection signals across six detectors that share a CLIP ViT-L/14 backbone finds that raw $|\Delta \text{logit}|$ under median-3 purification, applied through the EFFORT detector, separates adversarial inputs from clean inputs with AUROC 0.81-0.98.
Junghyun Kim, Seunghyun Kim, Ji-myung Woo
· 0 citations