paper-with-me

Papers

Speech Pattern based Black-box Model Watermarking for Automatic Speech Recognition

2021-10-19 · Haozhe Chen, Weiming Zhang, Kunlin Liu, Kejiang Chen, Han Fang, Nenghai Yu

As an effective method for intellectual property (IP) protection, model watermarking technology has been applied on a wide variety of deep neural networks (DNN), including speech classification models. However, how to design a black-box watermarking scheme for automatic speech recognition (ASR) models is still an unsolved problem, which is a significant demand for protecting remote ASR Application Programming Interface (API) deployed in cloud servers. Due to conditional independence assumption and label-detection-based evasion attack risk of ASR models, the black-box model watermarking scheme for speech classification models cannot apply to ASR models. In this paper, we propose the first black-box model watermarking framework for protecting the IP of ASR models. Specifically, we synthesize trigger audios by spreading the speech clips of model owners over the entire input audios and labeling the trigger audios with the stego texts, which hides the authorship information with linguistic steganography. Experiments on the state-of-the-art open-source ASR system DeepSpeech demonstrate the feasibility of the proposed watermarking scheme, which is robust against five kinds of attacks and has little impact on accuracy.

📄 PDF Abstract BibTeX arXiv:2110.09814

Code (0)

등록된 구현이 없습니다.

Tasks

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Linguistic steganographyspeech-recognitionSpeech Recognition

Similar Papers 제목 키워드 기반

Audio Codec Augmentation for Robust Collaborative Watermarking of Speech Synthesis

2024-09-20 · Lauri Juvela, Xin Wang

Automatic detection of synthetic speech is becoming increasingly important as current synthesis methods are both near indistinguishable from human speech and widely accessible to the public. Audio watermarking and other …

Face SwappingSpeech Synthesis

AudioMarkBench: Benchmarking Robustness of Audio Watermarking

2024-06-11 · Hongbin Liu, Moyang Guo, Zhengyuan Jiang, Lun Wang 외

The increasing realism of synthetic speech, driven by advancements in text-to-speech models, raises ethical concerns regarding impersonation and disinformation. Audio watermarking offers a promising solution via embeddin…

Benchmarkingtext-to-speechText to Speech

Collaborative Watermarking for Adversarial Speech Synthesis

2023-09-26 · Lauri Juvela, Xin Wang

Advances in neural speech synthesis have brought us technology that is not only close to human naturalness, but is also capable of instant voice cloning with little data, and is highly accessible with pre-trained models …

Speaker VerificationSpeech SynthesisSynthetic Speech DetectionVoice Cloning

The Watermark Shortcut: How Provenance Marking Sabotages Audio Deepfake Detection

2026-06-22 · Nicolas M. Müller, Pascal Debus arxiv

Provenance watermarking is increasingly treated as a safeguard for synthetic speech, whether built directly into speech-generation models such as Chatterbox, provided through dedicated techniques such as AudioSeal, or de…

Audio Deepfake Detection

SOLIDO: A Robust Watermarking Method for Speech Synthesis via Low-Rank Adaptation

2025-04-21 · Yue Li, Weizhi Liu, Dongdong Lin

The accelerated advancement of speech generative models has given rise to security issues, including model infringement and unauthorized abuse of content. Although existing generative watermarking techniques have propose…

parameter-efficient fine-tuningSpeech Synthesis