paper-with-me

홈 › Papers

Watermarking Language Models with Error Correcting Codes

2024-06-12 · Patrick Chao, Yan Sun, Edgar Dobriban, Hamed Hassani

Recent progress in large language models enables the creation of realistic machine-generated content. Watermarking is a promising approach to distinguish machine-generated text from human text, embedding statistical signals in the output that are ideally undetectable to humans. We propose a watermarking framework that encodes such signals through an error correcting code. Our method, termed robust binary code (RBC) watermark, introduces no noticeable degradation in quality. We evaluate our watermark on base and instruction fine-tuned models and find that our watermark is robust to edits, deletions, and translations. We provide an information-theoretic perspective on watermarking, a powerful statistical test for detection and for generating $p$-values, and theoretical guarantees. Our empirical findings suggest our watermark is fast, powerful, and robust, comparing favorably to the state-of-the-art.

📄 PDF Abstract BibTeX arXiv:2406.10281

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

BASE 설명 없음

Similar Papers 제목 키워드 기반

Pseudorandom Error-Correcting Codes

2024-02-14 · Miranda Christ, Sam Gunn

We construct pseudorandom error-correcting codes (or simply pseudorandom codes), which are error-correcting codes with the property that any polynomial number of codewords are pseudorandom to any computationally-bounded …

Robust and Verifiable Information Embedding Attacks to Deep Neural Networks via Error-Correcting Codes

2020-10-26 · Jinyuan Jia, Binghui Wang, Neil Zhenqiang Gong

In the era of deep learning, a user often leverages a third-party machine learning tool to train a deep neural network (DNN) classifier and then deploys the classifier as an end-user software product or a cloud service. …

TimeMark: A Trustworthy Time Watermarking Framework for Exact Generation-Time Recovery from AIGC

2026-04-14 · Shangkun Che, Silin Du, Ge Gao arxiv

The widespread use of Large Language Models (LLMs) in text generation has raised increasing concerns about intellectual property disputes. Watermarking techniques, which embed meta information into AI-generated content (…

Text Generation

Error-Correcting Codes For Approximate Neural Sequence Prediction

2021-11-16 · ACL ARR November 2021 11 · Anonymous

We propose a novel neural sequence prediction method based on \textit{error-correcting codes} that avoids exact softmax normalization and allows for a tradeoff between speed and performance. Error-correcting codes repres…

Language ModelingLanguage ModellingPredictionText Generation

UniMark: Unified Adaptive Multi-bit Watermarking for Autoregressive Image Generators

2026-04-12 · Yigit Yilmaz, Elena Petrova, Mehmet Kaya, Lucia Rossi 외 arxiv

Invisible watermarking for autoregressive (AR) image generation has recently gained attention as a means of protecting image ownership and tracing AI-generated content. However, existing approaches suffer from three key …

Semantic SimilarityImage Generation