paper-with-me

홈 › Papers

PMark: Towards Robust and Distortion-free Semantic-level Watermarking with Channel Constraints

2025-09-25 · Jiahao Huo, Shuliang Liu, Bin Wang, Junyan Zhang, Yibo Yan, Aiwei Liu, Xuming Hu, Mingxun Zhou arxiv

Semantic-level watermarking (SWM) for large language models (LLMs) enhances watermarking robustness against text modifications and paraphrasing attacks by treating the sentence as the fundamental unit. However, existing methods still lack strong theoretical guarantees of robustness, and reject-sampling-based generation often introduces significant distribution distortions compared with unwatermarked outputs. In this work, we introduce a new theoretical framework on SWM through the concept of proxy functions (PFs) $\unicode{x2013}$ functions that map sentences to scalar values. Building on this framework, we propose PMark, a simple yet powerful SWM method that estimates the PF median for the next sentence dynamically through sampling while enforcing multiple PF constraints (which we call channels) to strengthen watermark evidence. Equipped with solid theoretical guarantees, PMark achieves the desired distortion-free property and improves the robustness against paraphrasing-style attacks. We also provide an empirically optimized version that further removes the requirement for dynamical median estimation for better sampling efficiency. Experimental results show that PMark consistently outperforms existing SWM baselines in both text quality and robustness, offering a more effective paradigm for detecting machine-generated text. Our code will be released at this URL.

📄 PDF Abstract BibTeX arXiv:2509.21057

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

SepMark: Deep Separable Watermarking for Unified Source Tracing and Deepfake Detection

2023-05-10 · Xiaoshuai Wu, Xin Liao, Bo Ou

Malicious Deepfakes have led to a sharp conflict over distinguishing between genuine and forged faces. Although many countermeasures have been developed to detect Deepfakes ex-post, undoubtedly, passive forensics has not…

AttributeDecoderDeepFake DetectionFace Swapping

A Resilient and Accessible Distribution-Preserving Watermark for Large Language Models

2023-10-11 · Yihan Wu, Zhengmian Hu, Junfeng Guo, Hongyang Zhang 외

Watermarking techniques offer a promising way to identify machine-generated content via embedding covert information into the contents generated from language models. A challenge in the domain lies in preserving the dist…

Language ModelingLanguage Modelling

PASA: A Principled Embedding-Space Watermarking Approach for LLM-Generated Text under Semantic-Invariant Attacks

2026-05-09 · Zhenxin Ai, Haiyun He arxiv

Watermarking for large language models (LLMs) is a promising approach for detecting LLM-generated text and enabling responsible deployment. However, existing watermarking methods are often vulnerable to semantic-invarian…

Watermarking Vision-Language Pre-trained Models for Multi-modal Embedding as a Service

2023-11-10 · Yuanmin Tang, Jing Yu, Keke Gai, Xiangyan Qu 외

Recent advances in vision-language pre-trained models (VLPs) have significantly increased visual understanding and cross-modal analysis capabilities. Companies have emerged to provide multi-modal Embedding as a Service (…

Model extraction

CompMarkGS: Robust Watermarking for Compressed 3D Gaussian Splatting

2025-03-17 · Sumin In, Youngdong Jang, Utae Jeong, MinHyuk Jang 외

3D Gaussian Splatting (3DGS) enables rapid differentiable rendering for 3D reconstruction and novel view synthesis, leading to its widespread commercial use. Consequently, copyright protection via watermarking has become…

3DGS3D ReconstructionModel CompressionNovel View Synthesis+1