This paper carefully analyzes two state-of-the-art detection methods and their dimensionality reductions for convolutional layers and develops a novel reduction method with a controllable high-compression level.
Abstract
Despite the success of convolutional neural networks in image classification tasks and their general application in multi-modal models, their susceptibility to out-of-distribution and adversarial attack samples raises concerns regarding trustworthiness and safety. Among the approaches to tackle such issues, detection methods that analyze the model's intermediate activations to estimate a confidence score are a promising family that evaluates the decision process, relying on a dimensionality reduction step to enable efficient downstream processing of the high-dimensional activations. However, when considering convolutional layers, the dimensionality reduction methods in the literature either lack a mechanism to control the compression/information-loss trade-off or yield large representations. In this paper, we carefully analyze two state-of-the-art detection methods and their dimensionality reductions for convolutional layers and develop a novel reduction method with a controllable high-compression level. We extend these two state-of-the-art detection methods, enabling the usage of any dimensionality reduction, and evaluate their performance on out-of-distribution and adversarial attack detection. Results show that the detection methods with the proposed dimensionality reduction consistently perform better than, or comparable to, the strongest alternative. Furthermore, the proposed method is shown to reduce computation and memory footprints, given that it has the highest compression among the compared methods.
The rapid advancements in generative adversarial networks (GANs) have led to the production of highly realistic synthetic images, posing severe threats to the credibility and authenticity of digital media across social platforms, news outlets, and official documents. Passive detection methods tackle this problem by ide...
Deep Neural Networks (DNNs) remain vulnerable to adversarial perturbations, raising significant concerns in image processing applications, particularly in high-stakes domains such as medical imaging and security-critical systems. Most existing defense strategies are limited by domain specificity, architectural dependen...
Syamantak Sarkar, Nirmal Joseph, Sudhish N. George et al.· IEEE Transactions on Image P...· 0 citations
The integration of deep learning models into image preprocessing pipelines such as super-resolution introduces a largely unexplored attack vector for adversaries targeting downstream tasks. To ensure trustworthiness of critical imaging pipelines, we must be able to detect adversarial behavior within preprocessing model...
Emma J. Reid, Haley Duba-Sullivan, Tony G. Allen· 0 citations
The rapid advancement of deep generative models, especially Generative Adversarial Networks (GANs) and Diffusion Models, has escalated the creation of highly realistic synthetic media, posing significant threats to information security through misinformation and fraud. The core of current detection methodologies encomp...
Tian-Hua Tang· ITM Web of Conferences· 0 citations
This paper introduces a Feature-Space Ensemble Defense (FSED) framework, which involves adversarial training, joint confidence calibration, and class-conditional Mahalanobis feature-space anomaly scoring to facilitate powerful adversarial detection. The proposed method considers the final-layer uncertainty. It also mod...
Aliza Saadi, Vanya Shafiq, Aaleen Zainab et al.· 2026 IEEE International Conf...· 0 citations
Deepfake technology has advanced swiftly, enabling the rapid production of hyper-realistic synthetic media that pose considerable threats to digital security, privacy, military operations, and information integrity. This paper extensively examines visual intelligence and computer vision methodologies for deepfake detec...
Alexandros Gazis, Stylianos Pappas, T. Vavouras et al.· ICCK Transactions on Sensing...· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.