Skip to content
Preprint

Learning Unified Video and Image Representation for Video Face Forgery Detection

Aug 2026 · 0 citations · 54 references
Computer Science

TL;DR

A novel framework, UVIF, that utilizes additional annotated images to provide fine-grained supervision for detecting partial forgeries in videos, which outperforms state-of-theart methods in detecting partially forged videos while introducing no additional computational overhead is proposed.

Abstract

Face forgery detection is crucial for preserving the security and integrity of facial data given the rapid developments in face manipulation techniques and deep generative models. Existing methods for video face forgery detection typically assume that all frames in a forged video are manipulated, while detecting partially forged videos that contain only a subset of altered frames remains challenging. To address this issue, we propose a novel framework, UVIF, that utilizes additional annotated images to provide fine-grained supervision for detecting partial forgeries in videos. UVIF employs a unified encoder and a multi-task learning paradigm to jointly model facial videos and images for boosted video face forgery detection. A 2D backbone with temporal fusion modules is employed as the unified encoder. A pseudo labeling process is designed for video frames to bridge their representations with those of static images. A video-oriented feature alignment strategy is further introduced to reduce the distribution gap between videos and images. Extensive experiments on benchmark datasets demonstrate the effectiveness of our framework, which outperforms state-of-theart methods in detecting partially forged videos while introducing no additional computational overhead. Our code is available at https://github.com/haotianll/UVIF.

View source

Similar papers

Preprint Aug 2026

V-FIND: Revealing the Intrinsic Forgery Knowledge Encoded in Video Forgery Detectors

As generated videos become increasingly realistic, reliable video forgery detection is increasingly important. Existing studies typically optimize and use video forgery detectors as black boxes, while the latent forgery-discriminative knowledge inside them remains largely unexplored. Instead of continuing to rely on re...

Shi-Chao Kan, Chengpeng Hong, Jingtong Dou et al. · 0 citations
Conference Aug 2026

Explainable Deep Learning Framework for Accurate Detection and Interpretation of Copy-Move Forgery in Digital Images

In the era of advanced digital image generation techniques such as copy-move forgery, the NDV is having increasing concerns regarding image authenticity particularly with the spread of digital images through social media, media and courts. This type of forgery, where a part of an image is replicated and pasted on the s...

Shaheena K. V., D. S · 0 citations
Aug 2026

Image Forgery Detection Based on Fusion of Lightweight Deep Learning Models

A fusion-based lightweight deep learning framework for copy-move image forgery detection and localization that offers an efficient and practical solution for digital image authentication and is applicable to digital forensics, journalism, law enforcement, cyber security, and multimedia content verification.

K. Sumalini, K. B. Maruthiram · 0 citations
#machine learning Preprint Sep 2026

Image Classifiers are Efficient Self-Supervised Video Representation Learners

We introduce VideoMSN, a Masked Siamese Network framework for efficient self-supervised spatio-temporal representation learning in videos. Instead of relying on heavy 3D architectures or reconstruction-based autoencoders for learning with unlabeled data, we repurpose standard image Vision Transformers by representing v...

Owais Iqbal, Sudipta Sarkar, Shyam Marjit et al. · 0 citations
Review Open access 2026

Deep Learning-Based Face Detection, Feature Extraction, and Face Recognition from Video: A Comprehensive Review

A comparative analysis of existing studies is presented to highlight the evolution of deep learning techniques and their effectiveness in improving recognition accuracy and computational efficiency and emerging research directions are outlined to provide insights for future research.

Patel Bhautika Ronak · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.