The Band-Attention Modulation Network (BAM-Net) is proposed, a novel framework that pioneers learnable, fine-grained modulation of frequency components for forgery detection and exhibits exceptional generalization in cross-dataset, cross-compression, and cross-manipulation scenarios.
Abstract
Face forgery detection faces critical challenges in generalizing to unseen manipulation techniques and remaining robust under image compression, which often obscures subtle artifacts. Existing methods typically rely on fixed filters or coarse band separation, lacking the adaptability to learn task-specific spectral cues. To address this, we propose the Band-Attention Modulation Network (BAM-Net), a novel framework that pioneers learnable, fine-grained modulation of frequency components for forgery detection. At its core is the Band-Attention Modulation (BAM) mechanism, which transforms an image into its Discrete Cosine Transform (DCT) spectrogram and learns to dynamically reweight frequency bands along anti-diagonals. This process effectively enhances forgery-related spectral signatures while suppressing less informative ones, simulating an adaptive"inverse compression"that counters information loss. The modulated frequency information is then fused with the spatial domain to guide a lightweight yet effective spatial backbone equipped with distance-decayed attention for comprehensive feature extraction. Extensive experiments on FaceForensics++, Celeb-DF, and DFDC datasets demonstrate that BAM-Net achieves state-of-the-art performance. More importantly, it exhibits exceptional generalization in cross-dataset, cross-compression, and cross-manipulation scenarios, underscoring the vital role of adaptive frequency band modulation in building robust forgery detectors.
LGF-Net is proposed, a unified framework that jointly models spatial semantics and adaptive spectral cues for deepfake detection and achieves competitive intra-dataset and cross-dataset performance compared with several state-of-the-art deepfake detection methods.
Face forgery detectors often achieve strong results on controlled benchmarks, but their reliability under realistic image degradations remains limited. This paper presents a standardized benchmark for face forgery detection using the Multi-Dimensional Face Forgery Image (MFFI) dataset and evaluates performance on both...
L. Cunha, Lucas Sotomaior, Lucas Gasperin et al.· 0 citations
Deepfake technology, powered by deep learning models, enables the synthesis of highly realistic facial images and videos. However, in recent years, the misuse of deepfakes has posed severe challenges to both individual privacy and social trust. Consequently, this paper systematically reviews research pertaining to deep...
The increasing realism of image manipulations poses significant challenges for forgery localization. However, existing methods are hindered by the limited adaptability of constrained frequency filters and the dilution of subtle forensic cues in deep networks. To address these challenges, we propose the Laplacian pyrami...
Zhuo-Fei Liu, Wen-Jie Li, Yang Yu· IEEE Signal Processing Lette...· 0 citations
A generalizable deepfake detection framework that combines frequency-domain enhancement with feature disentanglement to encourage effective feature disentanglement and improve the discriminability of the learned forgery features is proposed.
Artificial intelligence-generated content (AIGC) has significantly improved the realism of manipulated facial images and videos, creating serious risks for social-network misinformation, content security, and digital trust. This study focuses on visual deepfake image detection rather than multimodal misinformation dete...
Chao-Ran Li· Journal of Cyber Security an...· 0 citations
Known for his clear and elegant writing style, Bertsekas shaped fields from control and optimization to large-scale computation and artificial intelligence.