The Impact of Audio Watermarking on Audio Anti-Spoofing Countermeasures
This paper presents the first study on the impact of audio watermarking on spoofing countermeasures. While anti-spoofing systems are essential for securing speech-based applications, the influence of widely used audio watermarking, originally designed for copyright protection, remains largely unexplored. We construct watermark-augmented training and evaluation datasets, named the Watermark-Spoofing dataset, by applying diverse handcrafted and neural watermarking methods to existing anti-spoofing datasets. Experiments show that watermarking consistently degrades anti-spoofing performance, with higher watermark density correlating with higher Equal Error Rates (EERs). To mitigate this, we propose the Knowledge-Preserving Watermark Learning (KPWL) framework, enabling models to adapt to watermark-induced shifts while preserving their original-domain spoofing detection capability. These findings reveal audio watermarking as a previously overlooked domain shift and establish the first benchmark for developing watermark-resilient anti-spoofing systems. All related protocols are publicly available at https://github.com/Alphawarheads/Watermark_Spoofing.git
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
What You Read Isn't What You Hear: Linguistic Sensitivity in Deepfake Speech Detection
Recent advances in text-to-speech technologies have enabled realistic voice generation, fueling audio-based deepfake attacks such as fraud and impersonation. While audio anti-spoofing systems are critical for detecting s…
Face SwappingSensitivitytext-to-speechText to SpeechSource Tracing of Audio Deepfake Systems
Recent progress in generative AI technology has made audio deepfakes remarkably more realistic. While current research on anti-spoofing systems primarily focuses on assessing whether a given audio sample is fake or genui…
Face Swappingtext-to-speechText to SpeechVoice ConversionA Deep Learning-based Audio-in-Image Watermarking Scheme
This paper presents a deep learning-based audio-in-image watermarking scheme. Audio-in-image watermarking is the process of covertly embedding and extracting audio watermarks on a cover-image. Using audio watermarks can …
Deep LearningAudioMarkBench: Benchmarking Robustness of Audio Watermarking
The increasing realism of synthetic speech, driven by advancements in text-to-speech models, raises ethical concerns regarding impersonation and disinformation. Audio watermarking offers a promising solution via embeddin…
Benchmarkingtext-to-speechText to SpeechComplex-valued neural networks for voice anti-spoofing
Current anti-spoofing and audio deepfake detection systems use either magnitude spectrogram-based features (such as CQT or Melspectrograms) or raw audio processed through convolution or sinc-layers. Both methods have dra…
Audio Deepfake DetectionDeepFake DetectionFace SwappingVoice Anti-spoofing