paper-with-me

홈 › Papers

Robustness Assessment and Enhancement of Text Watermarking for Google's SynthID

2025-08-27 · Xia Han, Qi Li, Jianbing Ni, Mohammad Zulkernine arxiv

Recent advances in LLM watermarking methods such as SynthID-Text by Google DeepMind offer promising solutions for tracing the provenance of AI-generated text. However, our robustness assessment reveals that SynthID-Text is vulnerable to meaning-preserving attacks, such as paraphrasing, copy-paste modifications, and back-translation, which can significantly degrade watermark detectability. To address these limitations, we propose SynGuard, a hybrid framework that combines the semantic alignment strength of Semantic Information Retrieval (SIR) with the probabilistic watermarking mechanism of SynthID-Text. Our approach jointly embeds watermarks at both lexical and semantic levels, enabling robust provenance tracking while preserving the original meaning. Experimental results across multiple attack scenarios show that SynGuard improves watermark recovery by an average of 11.1\% in F1 score compared to SynthID-Text. These findings demonstrate the effectiveness of semantic-aware watermarking in resisting real-world tampering. All code, datasets, and evaluation scripts are publicly available at: https://github.com/githshine/SynGuard.

📄 PDF Abstract BibTeX arXiv:2508.20228

Code (0)

등록된 구현이 없습니다.

Tasks

Information Retrieval

Similar Papers 제목 키워드 기반

On Google's SynthID-Text LLM Watermarking System: Theoretical Analysis and Empirical Validation

2026-03-03 · Romina Omidi, Yun Dong, Binghui Wang arxiv

Google's SynthID-Text, the first ever production-ready generative watermark system for large language model, designs a novel Tournament-based method that achieves the state-of-the-art detectability for identifying AI-gen…

The Brittleness of AI-Generated Image Watermarking Techniques: Examining Their Robustness Against Visual Paraphrasing Attacks

2024-08-19 · Niyar R Barman, Krish Sharma, Ashhar Aziz, Shashwat Bajpai 외

The rapid advancement of text-to-image generation systems, exemplified by models like Stable Diffusion, Midjourney, Imagen, and DALL-E, has heightened concerns about their potential misuse. In response, companies like Me…

DenoisingImage CaptioningImage GenerationText to Image Generation+1

Topic-Based Watermarks for Large Language Models

2024-04-02 · Alexander Nemecek, Yuzhou Jiang, Erman Ayday

The indistinguishability of Large Language Model (LLM) output from human-authored content poses significant challenges, raising concerns about potential misuse of AI-generated text and its influence on future AI model tr…

Language ModelingLanguage ModellingLarge Language ModelText Generation

RingID: Rethinking Tree-Ring Watermarking for Enhanced Multi-Key Identification

2024-04-22 · Hai Ci, Pei Yang, Yiren Song, Mike Zheng Shou

We revisit Tree-Ring Watermarking, a recent diffusion model watermarking method that demonstrates great robustness to various attacks. We conduct an in-depth study on it and reveal that the distribution shift unintention…

Watermark under Fire: A Robustness Evaluation of LLM Watermarking

2024-11-20 · Jiacheng Liang, Zian Wang, Lauren Hong, Shouling Ji 외

Various watermarking methods (``watermarkers'') have been proposed to identify LLM-generated texts; yet, due to the lack of unified evaluation platforms, many critical questions remain under-explored: i) What are the str…

Language ModelingLanguage Modellingmodel