Aug 2026· Journal of King Saud University: Computer and Information Sciences· Vol 38· 0 citations· 76 references
TL;DR
The proposed DSCM-FAS is a novel framework that combines inter-sample and intra-sample semantic consistency to enhance robustness and cross-domain generalization, without requiring auxiliary supervision, and demonstrates that DSCM-FAS consistently outperforms state-of-the-art methods under various protocols.
Abstract
Face recognition has been widely applied in identity authentication systems, but its vulnerability to various presentation attacks poses significant security risks. As a result, face anti-spoofing (FAS) has become one of the key technologies for ensuring the reliability of such systems. Most existing domain- generalized FAS (DGFAS) approaches rely on domain distribution alignment to learn cross-domain invariant representations. However, many of these methods independently model each sample while overlooking semantic consistency among cross-domain samples and different regions within each sample, making them vulnerable to semantic drift. Moreover, some existing methods introduce additional pixel-level supervision signals, such as pseudo-depth maps and binary masks, which increase annotation costs and limit their generalizability. To address these challenges, we propose DSCM-FAS, a novel framework that combines inter-sample and intra-sample semantic consistency to enhance robustness and cross-domain generalization, without requiring auxiliary supervision. Specifically, we design a Dual Semantic Consistency Module (DSCM). Across samples, contrastive learning is leveraged to learn discriminative representations, followed by the adaptive construction of a high-confidence similarity adjacency graph via statistical thresholding. The resulting graph is then fed into a graph convolutional network (GCN) to strengthen cross-domain semantic consistency, thereby improving the model's robustness to domain shifts. Within each sample, the intermediate embedding features are first partitioned into multiple patches, and Laplacian Regularization is introduced to constrain the semantic relationships among different patches. This effectively suppresses local noise interference and promotes the learning of more robust and semantically consistent feature representations. Extensive experiments on four public FAS datasets demonstrate that DSCM-FAS consistently outperforms state-of-the-art methods under various protocols. These results validate the effectiveness of our DSCM-FAS in improving the cross-domain generalization of FAS models.
The Semantic-Anchored Test-Time Domain Generalization (SA-TTDG) framework is introduced, introducing a Text-Anchored Style Projection (TASP), which utilizes rich linguistic priors from Vision-Language Models (VLMs) to initialize and strongly constrain learnable style bases.
Xiaosong Chang, Liang Shi, Ao Zhang· Computers, Materials & C...· 0 citations
This work proposes a robust, generalizable proactive face-swapping defense via semantic gradient divergence (SGD-Guard), and introduces an integrated feature gallery that uses CLIP features and a generalized identity feature, obtained by iteratively refining heterogeneous identity features into a homogeneous representa...
Seung-hyeok Back, Do Hyun Ki, Juwan Kim et al.· Proceedings of the Thirty-Fi...· 0 citations
Face Anti-Spoofing (FAS) is a crucial task for securing face recognition systems, yet its cross-domain generalization remains challenging. Recently, vision-language methods built upon pretrained models such as CLIP have shown promising performance in addressing these cross-domain scenarios. Nevertheless, existing appro...
Xiao-Meng Wei, Wen-Zhong Yang, Ya-Bo Yin et al.· Journal of King Saud Univers...· 0 citations
A few-shot, high-quality semantic annotation paradigm is effective for building trustworthy, explainable, and cost-efficient UFAD systems, and validating the robustness of semantic anchoring compared to models trained on massive short-text data.
Xiao-Yong Yu, Rong-Zhen Li, Shu-Ming Shi et al.· 0 citations
Face recognition has brought tremendous convenience to daily life, enabling applications like access control, mobile payment, and criminal identification. However, facial templates stored on servers pose significant privacy risks due to potential security breaches. To underscore this concern, facial template inversion-...
As image generation and editing technologies have progressed substantially, facial forgeries pose significant challenges to privacy and public safety. Due to limited ability to capture forgery cues, existing small-scale forgery detection models often struggle to generalize across various domains and unseen manipulation...
Feng-Ming Gu, Ming-Jie He, Zong-Hui Guo et al.· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.