paper-with-me

Papers

Investigating Detection and Obfuscation of Prompt Injection Attacks Against Software Reverse Engineering AI Agents

2026-05-29 · Brian Crawford, Patrick McClure arxiv

Agentic software reverse engineering systems are vulnerable to prompt injection attacks placed into the source code of executable binary files. This research demonstrates defensive tactics for detecting the presences of prompt injection strings in the decompiler output of adversarial example programs. Methods for obfuscating these attacks and subsequent methods for defending against these obfuscations are also explored. This research advances the understanding of risk and security of agentic software analysis systems necessary for their deployment into production-level cyber workflows.

📄 PDF Abstract BibTeX arXiv:2605.30677

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

A Multi-Agent LLM Defense Pipeline Against Prompt Injection Attacks

2025-09-16 · S M Asif Hossain, Ruksat Khan Shayoni, Mohd Ruhul Ameen, Akif Islam 외 arxiv

Prompt injection attacks represent a major vulnerability in Large Language Model (LLM) deployments, where malicious instructions embedded in user inputs can override system prompts and induce unintended behaviors. This p…

RIPA: Sensory-Vector Prompt Injection Attacks on LLM-Controlled ROS 2 Robots

2026-06-26 · Nima Dorzhiev arxiv

We present RIPA, the first systematic multi-channel empirical study of prompt injection attacks delivered through the sensory pipeline of a ROS 2-based LLM-controlled robotic system. Across 100 independent runs per injec…

Is Your Prompt Safe? Investigating Prompt Injection Attacks Against Open-Source LLMs

2025-05-20 · Jiawen Wang, Pritha Gupta, Ivan Habernal, Eyke Hüllermeier

Recent studies demonstrate that Large Language Models (LLMs) are vulnerable to different prompt-based attacks, generating harmful content or sensitive information. Both closed-source and open-source LLMs are underinvesti…

DataSentinel: A Game-Theoretic Detection of Prompt Injection Attacks

2025-04-15 · Yupei Liu, Yuqi Jia, Jinyuan Jia, Dawn Song 외

LLM-integrated applications and agents are vulnerable to prompt injection attacks, where an attacker injects prompts into their inputs to induce attacker-desired outputs. A detection method aims to determine whether a gi…

UniGuardian: A Unified Defense for Detecting Prompt Injection, Backdoor Attacks and Adversarial Attacks in Large Language Models

2025-02-18 · Huawei Lin, Yingjie Lao, Tong Geng, Tan Yu 외

Large Language Models (LLMs) are vulnerable to attacks like prompt injection, backdoor attacks, and adversarial attacks, which manipulate prompts or models to generate harmful outputs. In this paper, departing from tradi…

Text Generation