paper-with-me

홈 › Papers

Watermarking Needs Input Repetition Masking

2025-04-16 · David Khachaturov, Robert Mullins, Ilia Shumailov, Sumanth Dathathri

Recent advancements in Large Language Models (LLMs) raised concerns over potential misuse, such as for spreading misinformation. In response two counter measures emerged: machine learning-based detectors that predict if text is synthetic, and LLM watermarking, which subtly marks generated text for identification and attribution. Meanwhile, humans are known to adjust language to their conversational partners both syntactically and lexically. By implication, it is possible that humans or unwatermarked LLMs could unintentionally mimic properties of LLM generated text, making counter measures unreliable. In this work we investigate the extent to which such conversational adaptation happens. We call the concept $\textit{mimicry}$ and demonstrate that both humans and LLMs end up mimicking, including the watermarking signal even in seemingly improbable settings. This challenges current academic assumptions and suggests that for long-term watermarking to be reliable, the likelihood of false positives needs to be significantly lower, while longer word sequences should be used for seeding watermarking mechanisms.

📄 PDF Abstract BibTeX arXiv:2504.12229

Code (0)

등록된 구현이 없습니다.

Tasks

Misinformation

Similar Papers 제목 키워드 기반

Of-SemWat: High-payload text embedding for semantic watermarking of AI-generated images with arbitrary size

2025-09-29 · Benedetta Tondi, Andrea Costanzo, Mauro Barni arxiv

We propose a high-payload image watermarking method for textual embedding, where a semantic description of the image - which may also correspond to the input text prompt-, is embedded inside the image. In order to be abl…

dgMARK: Decoding-Guided Watermarking for Diffusion Language Models

2026-01-30 · Pyo Min Hong, Albert No arxiv

We propose dgMARK, a decoding-guided watermarking method for discrete diffusion language models (dLLMs). Unlike autoregressive models, dLLMs can generate tokens in arbitrary order. While an ideal conditional predictor wo…

RIGA: Covert and Robust White-Box Watermarking of Deep Neural Networks

2019-10-31 · Tianhao Wang, Florian Kerschbaum

Watermarking of deep neural networks (DNN) can enable their tracing once released by a data owner. In this paper, we generalize white-box watermarking algorithms for DNNs, where the data owner needs white-box access to t…

Inference Attack

Noise-to-mask Ratio Loss for Deep Neural Network based Audio Watermarking

2024-08-28 · Martin Moritz, Toni Olán, Tuomas Virtanen

Digital audio watermarking consists in inserting a message into audio signals in a transparent way and can be used to allow automatic recognition of audio material and management of the copyrights. We propose a perceptua…

Management

XAttnMark: Learning Robust Audio Watermarking with Cross-Attention

2025-02-06 · Yixin Liu, Lie Lu, Jihui Jin, Lichao Sun 외

The rapid proliferation of generative audio synthesis and editing technologies has raised significant concerns about copyright infringement, data provenance, and the spread of misinformation through deepfake audio. Water…

Audio SynthesisFace SwappingMisinformation