paper-with-me

Papers

Reject Threshold Adaptation for Open-Set Model Attribution of Deepfake Audio

2024-12-02 · Xinrui Yan, Jiangyan Yi, JianHua Tao, Yujie Chen, Hao Gu, Guanjun Li, Junzuo Zhou, Yong Ren, Tao Xu

Open environment oriented open set model attribution of deepfake audio is an emerging research topic, aiming to identify the generation models of deepfake audio. Most previous work requires manually setting a rejection threshold for unknown classes to compare with predicted probabilities. However, models often overfit training instances and generate overly confident predictions. Moreover, thresholds that effectively distinguish unknown categories in the current dataset may not be suitable for identifying known and unknown categories in another data distribution. To address the issues, we propose a novel framework for open set model attribution of deepfake audio with rejection threshold adaptation (ReTA). Specifically, the reconstruction error learning module trains by combining the representation of system fingerprints with labels corresponding to either the target class or a randomly chosen other class label. This process generates matching and non-matching reconstructed samples, establishing the reconstruction error distributions for each class and laying the foundation for the reject threshold calculation module. The reject threshold calculation module utilizes gaussian probability estimation to fit the distributions of matching and non-matching reconstruction errors. It then computes adaptive reject thresholds for all classes through probability minimization criteria. The experimental results demonstrate the effectiveness of ReTA in improving the open set model attributes of deepfake audio.

📄 PDF Abstract BibTeX arXiv:2412.01425

Code (0)

등록된 구현이 없습니다.

Tasks

Face Swapping

Methods 이 논문이 사용한 방법론

SET Dynamic Sparse Training method where weight mask is updated randomly periodically

Similar Papers 제목 키워드 기반

Towards Reliable Audio Deepfake Attribution and Model Recognition: A Multi-Level Autoencoder-Based Framework

2025-08-04 · Andrea Di Pierno, Luca Guarnera, Dario Allegra, Sebastiano Battiato arxiv

The proliferation of audio deepfakes poses a growing threat to trust in digital communications. While detection methods have advanced, attributing audio deepfakes to their source models remains an underexplored yet cruci…

Audio Deepfake Detection

Audio Deepfake Attribution: An Initial Dataset and Investigation

2022-08-21 · Xinrui Yan, Jiangyan Yi, JianHua Tao, Jie Chen

The rapid progress of deep speech synthesis models has posed significant threats to society such as malicious manipulation of content. This has led to an increase in studies aimed at detecting so-called deepfake audio. H…

Audio GenerationBinary ClassificationFace SwappingSpeech Synthesis

Contrastive Pseudo Learning for Open-World DeepFake Attribution

2023-09-20 · ICCV 2023 1 · Zhimin Sun, Shen Chen, Taiping Yao, Bangjie Yin 외

The challenge in sourcing attribution for forgery faces has gained widespread attention due to the rapid development of generative techniques. While many recent works have taken essential steps on GAN-generated faces, mo…

DeepFake DetectionFace SwappingPseudo Label

Task-Adaptive Negative Envision for Few-Shot Open-Set Recognition

2020-12-24 · CVPR 2022 1 · Shiyuan Huang, Jiawei Ma, Guangxing Han, Shih-Fu Chang

We study the problem of few-shot open-set recognition (FSOR), which learns a recognition system capable of both fast adaptation to new classes with limited labeled examples and rejection of unknown negative samples. Trad…

Few-Shot LearningOpen Set Learning

STOPA: A Database of Systematic VariaTion Of DeePfake Audio for Open-Set Source Tracing and Attribution

2025-05-26 · Anton Firc, Manasi Chibber, Jagabandhu Mishra, Vishwanath Pratap Singh 외

A key research area in deepfake speech detection is source tracing - determining the origin of synthesised utterances. The approaches may involve identifying the acoustic model (AM), vocoder model (VM), or other generati…

DeepFake DetectionFace Swapping