paper-with-me

Papers

STOPA: A Database of Systematic VariaTion Of DeePfake Audio for Open-Set Source Tracing and Attribution

2025-05-26 · Anton Firc, Manasi Chibber, Jagabandhu Mishra, Vishwanath Pratap Singh, Tomi Kinnunen, Kamil Malinka

A key research area in deepfake speech detection is source tracing - determining the origin of synthesised utterances. The approaches may involve identifying the acoustic model (AM), vocoder model (VM), or other generation-specific parameters. However, progress is limited by the lack of a dedicated, systematically curated dataset. To address this, we introduce STOPA, a systematically varied and metadata-rich dataset for deepfake speech source tracing, covering 8 AMs, 6 VMs, and diverse parameter settings across 700k samples from 13 distinct synthesisers. Unlike existing datasets, which often feature limited variation or sparse metadata, STOPA provides a systematically controlled framework covering a broader range of generative factors, such as the choice of the vocoder model, acoustic model, or pretrained weights, ensuring higher attribution reliability. This control improves attribution accuracy, aiding forensic analysis, deepfake detection, and generative model transparency.

📄 PDF Abstract BibTeX arXiv:2505.19644

Code (0)

등록된 구현이 없습니다.

Tasks

DeepFake DetectionFace Swapping

Similar Papers 제목 키워드 기반

Vulnerability of Automatic Identity Recognition to Audio-Visual Deepfakes

2023-11-29 · Pavel Korshunov, Haolin Chen, Philip N. Garner, Sebastien Marcel

The task of deepfakes detection is far from being solved by speech or vision researchers. Several publicly available databases of fake synthetic video and speech were built to aid the development of detection methods. Ho…

Face RecognitionFace SwappingSpeaker Recognitiontext-to-speech+2

Heterogeneity over Homogeneity: Investigating Multilingual Speech Pre-Trained Models for Detecting Audio Deepfake

2024-03-31 · Orchid Chetia Phukan, Gautam Siddharth Kashyap, Arun Balaji Buduru, Rajesh Sharma

In this work, we investigate multilingual speech Pre-Trained models (PTMs) for Audio deepfake detection (ADD). We hypothesize that multilingual PTMs trained on large-scale diverse multilingual data gain knowledge about d…

Audio Deepfake DetectionDeepFake DetectionEmotion RecognitionFace Swapping

Towards Generalizable Deepfake Detection via Forgery-aware Audio-Visual Adaptation: A Variational Bayesian Approach

2025-11-24 · Fan Nie, Jiangqun Ni, Jian Zhang, Bin Zhang 외 arxiv

The widespread application of AIGC contents has brought not only unprecedented opportunities, but also potential security concerns, e.g., audio-visual deepfakes. Therefore, it is of great importance to develop an effecti…

DeepFake Detection

Zero-Shot to Zero-Lies: Detecting Bengali Deepfake Audio through Transfer Learning

2025-12-25 · Most. Sharmin Sultana Samu, Md. Rakibul Islam, Md. Zahid Hossain, Md. Kamrozzaman Bhuiyan 외 arxiv

The rapid growth of speech synthesis and voice conversion systems has made deepfake audio a major security concern. Bengali deepfake detection remains largely unexplored. In this work, we study automatic detection of Ben…

DeepFake DetectionTransfer LearningVoice ConversionSpeech Synthesis

Audio Deepfake Attribution: An Initial Dataset and Investigation

2022-08-21 · Xinrui Yan, Jiangyan Yi, JianHua Tao, Jie Chen

The rapid progress of deep speech synthesis models has posed significant threats to society such as malicious manipulation of content. This has led to an increase in studies aimed at detecting so-called deepfake audio. H…

Audio GenerationBinary ClassificationFace SwappingSpeech Synthesis