paper-with-me

홈 › Papers

A Content-Preserving Secure Linguistic Steganography

2025-11-16 · Lingyun Xiang, Chengfu Ou, Xu He, Zhongliang Yang, Yuling Liu arxiv

Existing linguistic steganography methods primarily rely on content transformations to conceal secret messages. However, they often cause subtle yet looking-innocent deviations between normal and stego texts, posing potential security risks in real-world applications. To address this challenge, we propose a content-preserving linguistic steganography paradigm for perfectly secure covert communication without modifying the cover text. Based on this paradigm, we introduce CLstega (\textit{C}ontent-preserving \textit{L}inguistic \textit{stega}nography), a novel method that embeds secret messages through controllable distribution transformation. CLstega first applies an augmented masking strategy to locate and mask embedding positions, where MLM(masked language model)-predicted probability distributions are easily adjustable for transformation. Subsequently, a dynamic distribution steganographic coding strategy is designed to encode secret messages by deriving target distributions from the original probability distributions. To achieve this transformation, CLstega elaborately selects target words for embedding positions as labels to construct a masked sentence dataset, which is used to fine-tune the original MLM, producing a target MLM capable of directly extracting secret messages from the cover text. This approach ensures perfect security of secret messages while fully preserving the integrity of the original cover text. Experimental results show that CLstega can achieve a 100\% extraction success rate, and outperforms existing methods in security, effectively balancing embedding capacity and security.

📄 PDF Abstract BibTeX arXiv:2511.12565

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Efficient Provably Secure Linguistic Steganography via Range Coding

2026-04-09 · Ruiyi Yan, Yugo Murawaki arxiv

Linguistic steganography involves embedding secret messages within seemingly innocuous texts to enable covert communication. Provable security, which is a long-standing goal and key motivation, has been extended to langu…

Provably Secure Disambiguating Neural Linguistic Steganography

2024-03-26 · Yuang Qi, Kejiang Chen, Kai Zeng, Weiming Zhang 외

Recent research in provably secure neural linguistic steganography has overlooked a crucial aspect: the sender must detokenize stegotexts to avoid raising suspicion from the eavesdropper. The segmentation ambiguity probl…

Linguistic steganography

A High-Capacity and Secure Disambiguation Algorithm for Neural Linguistic Steganography

2025-09-26 · Yapei Feng, Feng Jiang, Shanhao Wu, Hua Zhong arxiv

Neural linguistic steganography aims to embed information into natural text while preserving statistical undetectability. A fundamental challenge in this ffeld stems from tokenization ambiguity in modern tokenizers, whic…

A Secure and Disambiguating Approach for Generative Linguistic Steganography

2023-08-11 · IEEE Signal Processing Letters 2023 8 · Ruiyi Yan, Yating Yang, Tian Song

Segmentation ambiguity in generative linguistic steganography could induce decoding errors. One existing disambiguating way is removing the tokens whose mapping words are the prefixes of others in each candidate pool. Ho…

Linguistic steganographySegmentationSteganalysis

Frustratingly Easy Edit-based Linguistic Steganography with a Masked Language Model

2021-04-20 · NAACL 2021 4 · Honai Ueoka, Yugo Murawaki, Sadao Kurohashi

With advances in neural language models, the focus of linguistic steganography has shifted from edit-based approaches to generation-based ones. While the latter's payload capacity is impressive, generating genuine-lookin…

Language ModelingLanguage ModellingLinguistic steganography