paper-with-me

홈 › Papers

AT-ADD: All-Type Audio Deepfake Detection Challenge Evaluation Plan

2026-04-09 · Yuankun Xie, Haonan Cheng, Jiayi Zhou, Xiaoxuan Guo, Tao Wang, Jian Liu, Weiqiang Wang, Ruibo Fu, Xiaopeng Wang, Hengyan Huang, Xiaoying Huang, Long Ye, Guangtao Zhai arxiv

The rapid advancement of Audio Large Language Models (ALLMs) has enabled cost-effective, high-fidelity generation and manipulation of both speech and non-speech audio, including sound effects, singing voices, and music. While these capabilities foster creativity and content production, they also introduce significant security and trust challenges, as realistic audio deepfakes can now be generated and disseminated at scale. Existing audio deepfake detection (ADD) countermeasures (CMs) and benchmarks, however, remain largely speech-centric, often relying on speech-specific artifacts and exhibiting limited robustness to real-world distortions, as well as restricted generalization to heterogeneous audio types and emerging spoofing techniques. To address these gaps, we propose the All-Type Audio Deepfake Detection (AT-ADD) Grand Challenge for ACM Multimedia 2026, designed to bridge controlled academic evaluation with practical multimedia forensics. AT-ADD comprises two tracks: (1) Robust Speech Deepfake Detection, which evaluates detectors under real-world scenarios and against unseen, state-of-the-art speech generation methods; and (2) All-Type Audio Deepfake Detection, which extends detection beyond speech to diverse, unknown audio types and promotes type-agnostic generalization across speech, sound, singing, and music. By providing standardized datasets, rigorous evaluation protocols, and reproducible baselines, AT-ADD aims to accelerate the development of robust and generalizable audio forensic technologies, supporting secure communication, reliable media verification, and responsible governance in an era of pervasive synthetic audio.

📄 PDF Abstract BibTeX arXiv:2604.08184

Code (0)

등록된 구현이 없습니다.

Tasks

Audio Deepfake Detection

Similar Papers 제목 키워드 기반

What to Remember: Self-Adaptive Continual Learning for Audio Deepfake Detection

2023-12-15 · Xiaohui Zhang, Jiangyan Yi, Chenglong Wang, Chuyuan Zhang 외

The rapid evolution of speech synthesis and voice conversion has raised substantial concerns due to the potential misuse of such technology, prompting a pressing need for effective audio deepfake detection mechanisms. Ex…

Audio Deepfake DetectionContinual LearningDeepFake DetectionFace Swapping+2

WavLM model ensemble for audio deepfake detection

2024-08-14 · David Combei, Adriana Stan, Dan Oneata, Horia Cucu

Audio deepfake detection has become a pivotal task over the last couple of years, as many recent speech synthesis and voice cloning systems generate highly realistic speech samples, thus enabling their use in malicious a…

Audio Deepfake DetectionData AugmentationDeepFake DetectionFace Swapping+3

Detect All-Type Deepfake Audio: Wavelet Prompt Tuning for Enhanced Auditory Perception

2025-04-09 · Yuankun Xie, Ruibo Fu, Zhiyong Wang, Xiaopeng Wang 외

The rapid advancement of audio generation technologies has escalated the risks of malicious deepfake audio across speech, sound, singing voice, and music, threatening multimedia security and trust. While existing counter…

AllAudio Deepfake DetectionAudio GenerationDeepFake Detection+2

AV-Deepfake1M++: A Large-Scale Audio-Visual Deepfake Benchmark with Real-World Perturbations

2025-07-28 · Zhixi Cai, Kartik Kuckreja, Shreya Ghosh, Akanksha Chuchra 외 arxiv

The rapid surge of text-to-speech and face-voice reenactment models makes video fabrication easier and highly realistic. To encounter this problem, we require datasets that rich in type of generation methods and perturba…

Does Current Deepfake Audio Detection Model Effectively Detect ALM-based Deepfake Audio?

2024-08-20 · Yuankun Xie, Chenxu Xiong, Xiaopeng Wang, Zhiyong Wang 외

Currently, Audio Language Models (ALMs) are rapidly advancing due to the developments in large language models and audio neural codecs. These ALMs have significantly lowered the barrier to creating deepfake audio, genera…

Audio Deepfake DetectionDeepFake DetectionFace Swapping