paper-with-me

홈 › Papers

Tool Use Enables Undetectable Steganography in Multi-Agent LLM Systems

2026-06-25 · Jimmy Laurence Rippin, Simon C. Marshall, David Demitri Africa, Christian Schroeder de Witt arxiv

Increasingly autonomous agentic AI systems pose novel multi-agent risks, such as secret collusion via covert communication channels. The natural defence to these collusion attempts is to monitor plain-text communication, but the efficacy of monitors has been called into doubt by increasingly sophisticated model steganography; indeed, some theoretical schemes have been proposed that are information-theoretically or computationally indistinguishable from good-faith plain-text communication. In this paper, we demonstrate that the complexity of these schemes is no longer a safety barrier, as agentic coding models can already produce undetectable stegosystems when given realistic tool usage, such as code execution or accessing research papers through web searches. Agents also adapt when key ingredients are missing, for example, by adding model-sampling components or implementing related keyed coding schemes. We then frame tacit steganographic coordination between agents as a Schelling-point problem and introduce coordination metrics for estimating when two agents are likely to select compatible schemes without explicit prior agreement. Our results suggest a shift in the threat model for covert communication between AI agents, where the main barrier is no longer whether frontier agents can understand and implement sophisticated stegosystems, but coordination: whether independently acting agents can converge on compatible schemes, keys, and parameters. We find substantial convergence on broad scheme families but limited strict one-shot coordination, suggesting that shared artefacts, repeated interaction, and tool-mediated search are the settings where covert communication risks are most acute. Overall, our findings provide empirical grounding for the recent strategic confinement hypothesis, which assumes that capable agents can construct covert channels that survive monitoring.

📄 PDF Abstract BibTeX arXiv:2606.28425

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Pseudorandom Error-Correcting Codes

2024-02-14 · Miranda Christ, Sam Gunn

We construct pseudorandom error-correcting codes (or simply pseudorandom codes), which are error-correcting codes with the property that any polynomial number of codewords are pseudorandom to any computationally-bounded …

Hidden in Plain Text: Emergence & Mitigation of Steganographic Collusion in LLMs

2024-10-02 · Yohan Mathew, Ollie Matthews, Robert McCarthy, Joan Velja 외

The rapid proliferation of frontier model agents promises significant societal advances but also raises concerns about systemic risks arising from unsafe interactions. Collusion to the disadvantage of others has been ide…

In-Context Reinforcement Learningreinforcement-learningReinforcement Learning

IBRSteG: Learning a Generalizable Steganography Framework for 3D Gaussian Splatting

2026-06-29 · Fanye Kong, Hongyu Xia, Yu Zheng, Boyang Gong 외 arxiv

Recent advances in deep learning have notably improved steganographic message hiding. However, designing a generalizable steganographic approach for 3D Gaussian Splatting (3DGS) that can embed meaningful 3D scene content…

Steganographic Embeddings as an Effective Data Augmentation

2025-02-21 · Nicholas DiSalvo

Image Steganography is a cryptographic technique that embeds secret information into an image, ensuring the hidden data remains undetectable to the human eye while preserving the image's original visual integrity. Least …

Data Augmentationimage-classificationImage ClassificationImage Steganography

A Technical Review on Comparison and Estimation of Steganographic Tools

2025-08-26 · Ms. Preeti P. Bhatt, Rakesh R. Savant arxiv

Steganography is technique of hiding a data under cover media using different steganography tools. Image steganography is hiding of data (Text/Image/Audio/Video) under a cover as Image. This review paper presents classif…