paper-with-me

Papers

CodecFake: Enhancing Anti-Spoofing Models Against Deepfake Audios from Codec-Based Speech Synthesis Systems

2024-06-11 · Haibin Wu, Yuan Tseng, Hung-Yi Lee

Current state-of-the-art (SOTA) codec-based audio synthesis systems can mimic anyone's voice with just a 3-second sample from that specific unseen speaker. Unfortunately, malicious attackers may exploit these technologies, causing misuse and security issues. Anti-spoofing models have been developed to detect fake speech. However, the open question of whether current SOTA anti-spoofing models can effectively counter deepfake audios from codec-based speech synthesis systems remains unanswered. In this paper, we curate an extensive collection of contemporary SOTA codec models, employing them to re-create synthesized speech. This endeavor leads to the creation of CodecFake, the first codec-based deepfake audio dataset. Additionally, we verify that anti-spoofing models trained on commonly used datasets cannot detect synthesized speech from current codec-based speech generation systems. The proposed CodecFake dataset empowers these models to counter this challenge effectively.

📄 PDF Abstract BibTeX arXiv:2406.07237

Code (0)

등록된 구현이 없습니다.

Tasks

Audio SynthesisFace SwappingSpeech Synthesis

Similar Papers 제목 키워드 기반

Codecfake: An Initial Dataset for Detecting LLM-based Deepfake Audio

2024-06-12 · Yi Lu, Yuankun Xie, Ruibo Fu, Zhengqi Wen 외

With the proliferation of Large Language Model (LLM) based deepfake audio, there is an urgent need for effective detection methods. Previous deepfake audio generation methods typically involve a multi-step generation pro…

Audio Deepfake DetectionAudio GenerationDeepFake DetectionFace Swapping+3

The Codecfake Dataset and Countermeasures for the Universally Detection of Deepfake Audio

2024-05-08 · Yuankun Xie, Yi Lu, Ruibo Fu, Zhengqi Wen 외

With the proliferation of Audio Language Model (ALM) based deepfake audio, there is an urgent need for generalized detection methods. ALM-based deepfake audio currently exhibits widespread, high deception, and type versa…

Audio Deepfake DetectionAudio GenerationDeepFake DetectionFace Swapping+2

Towards Generalized Source Tracing for Codec-Based Deepfake Speech

2025-06-08 · Xuanjun Chen, I-Ming Lin, Lin Zhang, Haibin Wu 외

Recent attempts at source tracing for codec-based deepfake speech (CodecFake), generated by neural audio codec-based speech generation (CoSG) models, have exhibited suboptimal performance. However, how to train source tr…

Face Swapping

End-to-end Audio Deepfake Detection from RAW Waveforms: a RawNet-Based Approach with Cross-Dataset Evaluation

2025-04-29 · Andrea Di Pierno, Luca Guarnera, Dario Allegra, Sebastiano Battiato

Audio deepfakes represent a growing threat to digital security and trust, leveraging advanced generative models to produce synthetic speech that closely mimics real human voices. Detecting such manipulations is especiall…

Audio Deepfake DetectionDeepFake DetectionFace Swapping

Source Tracing of Audio Deepfake Systems

2024-07-10 · Nicholas Klein, Tianxiang Chen, Hemlata Tak, Ricardo Casal 외

Recent progress in generative AI technology has made audio deepfakes remarkably more realistic. While current research on anti-spoofing systems primarily focuses on assessing whether a given audio sample is fake or genui…

Face Swappingtext-to-speechText to SpeechVoice Conversion