paper-with-me

Papers

Audio Codec Augmentation for Robust Collaborative Watermarking of Speech Synthesis

2024-09-20 · Lauri Juvela, Xin Wang

Automatic detection of synthetic speech is becoming increasingly important as current synthesis methods are both near indistinguishable from human speech and widely accessible to the public. Audio watermarking and other active disclosure methods of are attracting research activity, as they can complement traditional deepfake defenses based on passive detection. In both active and passive detection, robustness is of major interest. Traditional audio watermarks are particularly susceptible to removal attacks by audio codec application. Most generated speech and audio content released into the wild passes through an audio codec purely as a distribution method. We recently proposed collaborative watermarking as method for making generated speech more easily detectable over a noisy but differentiable transmission channel. This paper extends the channel augmentation to work with non-differentiable traditional audio codecs and neural audio codecs and evaluates transferability and effect of codec bitrate over various configurations. The results show that collaborative watermarking can be reliably augmented by black-box audio codecs using a waveform-domain straight-through-estimator for gradient approximation. Furthermore, that results show that channel augmentation with a neural audio codec transfers well to traditional codecs. Listening tests demonstrate collaborative watermarking incurs negligible perceptual degradation with high bitrate codecs or DAC at 8kbps.

📄 PDF Abstract BibTeX arXiv:2409.13382

Code (1)

ljuvela/collaborative-watermarking-with-codecs 공식 구현 pytorch

Tasks

Face SwappingSpeech Synthesis

Methods 이 논문이 사용한 방법론

DAC 설명 없음

Similar Papers 제목 키워드 기반

Feature-Aligned Speech Watermarking for Robustness to Reconstruction Distortions

2026-06-10 · Haiyun Li, Shuhai Peng, Zhisheng Zhang, Jingran Xie 외 arxiv

Audio watermarking aims to embed identifiable information into audio while remaining imperceptible. Existing methods adopt high-fidelity, low-energy designs to preserve perceptual quality, but the resulting watermarks la…

Latent-Mark: An Audio Watermark Robust to Neural Codec Compression

2026-03-05 · Yen-Shan Chen, Shih-Yu Lai, Ying-Jung Tsou, Yi-Cheng Lin 외 arxiv

While existing audio watermarking techniques have achieved strong robustness against traditional digital signal processing (DSP) attacks, they remain vulnerable to neural compression. This occurs because modern neural au…

A Comprehensive Real-World Assessment of Audio Watermarking Algorithms: Will They Survive Neural Codecs?

2025-05-26 · Yigitcan Özer, Woosung Choi, Joan Serrà, Mayank Kumar Singh 외

We introduce the Robust Audio Watermarking Benchmark (RAW-Bench), a benchmark for evaluating deep learning-based audio watermarking methods with standardized and systematic comparisons. To simulate real-world usage, we i…

Baseline Systems For The 2025 Low-Resource Audio Codec Challenge

2025-09-30 · Yusuf Ziya Isik, Rafał Łaganowski arxiv

The Low-Resource Audio Codec (LRAC) Challenge aims to advance neural audio coding for deployment in resource-constrained environments. The first edition focuses on low-resource neural speech codecs that must operate reli…

Collaborative Watermarking for Adversarial Speech Synthesis

2023-09-26 · Lauri Juvela, Xin Wang

Advances in neural speech synthesis have brought us technology that is not only close to human naturalness, but is also capable of instant voice cloning with little data, and is highly accessible with pre-trained models …

Speaker VerificationSpeech SynthesisSynthetic Speech DetectionVoice Cloning