paper-with-me

홈 › Papers

CapST: An Enhanced and Lightweight Model Attribution Approach for Synthetic Videos

2023-11-07 · Wasim Ahmad, Yan-Tsung Peng, Yuan-Hao Chang, Gaddisa Olani Ganfure, Sarwar Khan, Sahibzada Adil Shahzad

Deepfake videos, generated through AI faceswapping techniques, have garnered considerable attention due to their potential for powerful impersonation attacks. While existing research primarily focuses on binary classification to discern between real and fake videos, however determining the specific generation model for a fake video is crucial for forensic investigation. Addressing this gap, this paper investigates the model attribution problem of Deepfake videos from a recently proposed dataset, Deepfakes from Different Models (DFDM), derived from various Autoencoder models. The dataset comprises 6,450 Deepfake videos generated by five distinct models with variations in encoder, decoder, intermediate layer, input resolution, and compression ratio. This study formulates Deepfakes model attribution as a multiclass classification task, proposing a segment of VGG19 as a feature extraction backbone, known for its effectiveness in imagerelated tasks, while integrated a Capsule Network with a Spatio-Temporal attention mechanism. The Capsule module captures intricate hierarchies among features for robust identification of deepfake attributes. Additionally, the video-level fusion technique leverages temporal attention mechanisms to handle concatenated feature vectors, capitalizing on inherent temporal dependencies in deepfake videos. By aggregating insights across frames, our model gains a comprehensive understanding of video content, resulting in more precise predictions. Experimental results on the deepfake benchmark dataset (DFDM) demonstrate the efficacy of our proposed method, achieving up to a 4% improvement in accurately categorizing deepfake videos compared to baseline models while demanding fewer computational resources.

📄 PDF Abstract BibTeX arXiv:2311.03782

Code (1)

wasim004/CapST 공식 구현 pytorch

Tasks

Decoder

Similar Papers 제목 키워드 기반

Capstan-driven Continuum Surgical Robot: Design, Modeling, and Perception

2026-08-13 · Gang Zhang, Yufu Qiu, Junyan Yan, Wenhui Zeng 외 arxiv

Shape and force sensing have long been critical bottlenecks in the development of compact capstan-driven continuum surgical robots, primarily due to the difficulty of obtaining cable tension information within the confin…

Pose Estimation

SAGA: Source Attribution of Generative AI Videos

2025-11-16 · Rohit Kundu, Vishal Mohanty, Hao Xiong, Shan Jia 외 arxiv

The proliferation of generative AI has led to hyper-realistic synthetic videos, escalating misuse risks and outstripping binary real/fake detectors. We introduce SAGA (Source Attribution of Generative AI videos), the fir…

Beyond Deepfake Images: Detecting AI-Generated Videos

2024-04-24 · Danial Samadi Vahdati, Tai D. Nguyen, Aref Azizpour, Matthew C. Stamm

Recent advances in generative AI have led to the development of techniques to generate visually realistic synthetic video. While a number of techniques have been developed to detect AI-generated synthetic images, in this…

Face SwappingFew-Shot Learning

FAME: A Lightweight Spatio-Temporal Network for Model Attribution of Face-Swap Deepfakes

2025-06-13 · Wasim Ahmad, Yan-Tsung Peng, Yuan-Hao Chang

The widespread emergence of face-swap Deepfake videos poses growing risks to digital security, privacy, and media integrity, necessitating effective forensic tools for identifying the source of such manipulations. Althou…

DeepFake DetectionFace Swapping

CapStARE: Capsule-based Sequential Architecture for Robust and Efficient Gaze Estimation

2025-09-24 · Miren Samaniego, Igor Rodriguez, Elena Lazkano arxiv

Human gaze estimation is essential for applications such as human-computer interaction, social robotics, and assistive systems. However, achieving accurate, interpretable, and real-time performance in unconstrained envir…

Computational EfficiencyGaze Estimation