paper-with-me

홈 › Papers

Less is More: Sparse Watermarking in LLMs with Enhanced Text Quality

2024-07-17 · Duy C. Hoang, Hung T. Q. Le, Rui Chu, Ping Li, Weijie Zhao, Yingjie Lao, Khoa D. Doan

With the widespread adoption of Large Language Models (LLMs), concerns about potential misuse have emerged. To this end, watermarking has been adapted to LLM, enabling a simple and effective way to detect and monitor generated text. However, while the existing methods can differentiate between watermarked and unwatermarked text with high accuracy, they often face a trade-off between the quality of the generated text and the effectiveness of the watermarking process. In this work, we present a novel type of LLM watermark, Sparse Watermark, which aims to mitigate this trade-off by applying watermarks to a small subset of generated tokens distributed across the text. The key strategy involves anchoring watermarked tokens to words that have specific Part-of-Speech (POS) tags. Our experimental results demonstrate that the proposed watermarking scheme achieves high detectability while generating text that outperforms previous LLM watermarking methods in quality across various tasks

📄 PDF Abstract BibTeX arXiv:2407.13803

Code (1)

mail-research/sparse-llm-watermarking 공식 구현 pytorch

Tasks

POS

Similar Papers 제목 키워드 기반

Mark Your LLM: Detecting the Misuse of Open-Source Large Language Models via Watermarking

2025-03-06 · Yijie Xu, Aiwei Liu, Xuming Hu, Lijie Wen 외

As open-source large language models (LLMs) like Llama3 become more capable, it is crucial to develop watermarking techniques to detect their potential misuse. Existing watermarking methods either add watermarks during L…

Video Watermarking: Safeguarding Your Video from (Unauthorized) Annotations by Video-based LLMs

2024-07-02 · Jinmin Li, Kuofeng Gao, Yang Bai, Jingyun Zhang 외

The advent of video-based Large Language Models (LLMs) has significantly enhanced video understanding. However, it has also raised some safety concerns regarding data protection, as videos can be more easily annotated, e…

Video Understanding

WatME: Towards Lossless Watermarking Through Lexical Redundancy

2023-11-16 · Liang Chen, Yatao Bian, Yang Deng, Deng Cai 외

Text watermarking has emerged as a pivotal technique for identifying machine-generated text. However, existing methods often rely on arbitrary vocabulary partitioning during decoding to embed watermarks, which compromise…

Instruction FollowingLanguage ModellingLogical ReasoningResponse Generation+1

RingID: Rethinking Tree-Ring Watermarking for Enhanced Multi-Key Identification

2024-04-22 · Hai Ci, Pei Yang, Yiren Song, Mike Zheng Shou

We revisit Tree-Ring Watermarking, a recent diffusion model watermarking method that demonstrates great robustness to various attacks. We conduct an in-depth study on it and reveal that the distribution shift unintention…

Token-Specific Watermarking with Enhanced Detectability and Semantic Coherence for Large Language Models

2024-02-28 · Mingjia Huo, Sai Ashish Somayajula, Youwei Liang, Ruisi Zhang 외

Large language models generate high-quality responses with potential misinformation, underscoring the need for regulation by distinguishing AI-generated and human-written texts. Watermarking is pivotal in this context, w…

Misinformation