paper-with-me

홈 › Papers

SABER: A Stealthy Agentic Black-Box Attack Framework for Vision-Language-Action Models

2026-03-26 · Xiyang Wu, Guangyao Shi, Qingzi Wang, Zongxia Li, Amrit Singh Bedi, Dinesh Manocha arxiv

Vision-language-action (VLA) models enable robots to follow natural-language instructions grounded in visual observations, but the instruction channel also introduces a critical vulnerability: small textual perturbations can alter downstream robot behavior. Systematic robustness evaluation therefore requires a black-box attacker that can generate minimal yet effective instruction edits across diverse VLA models. To this end, we present SABER, an agent-centric approach for automatically generating instruction-based adversarial attacks on VLA models under bounded edit budgets. SABER uses a GRPO-trained ReAct attacker to generate small, plausible adversarial instruction edits using character-, token-, and prompt-level tools under a bounded edit budget that induces targeted behavioral degradation, including task failure, unnecessarily long execution, and increased constraint violations. On the LIBERO benchmark across six state-of-the-art VLA models, SABER reduces task success by 20.6%, increases action-sequence length by 55%, and raises constraint violations by 33%, while requiring 21.1% fewer tool calls and 54.7% fewer character edits than strong GPT-based baselines. These results show that small, plausible instruction edits are sufficient to substantially degrade robot execution, and that an agentic black-box pipeline offers a practical, scalable, and adaptive approach for red-teaming robotic foundation models. The codebase is publicly available at https://github.com/wuxiyang1996/SABER.

📄 PDF Abstract BibTeX arXiv:2603.24935

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Black-box Stealthy GPS Attacks on Unmanned Aerial Vehicles

2024-09-17 · Amir Khazraei, Haocheng Meng, Miroslav Pajic

This work focuses on analyzing the vulnerability of unmanned aerial vehicles (UAVs) to stealthy black-box false data injection attacks on GPS measurements. We assume that the quadcopter is equipped with IMU and GPS senso…

Sensor Fusion

DrunkAgent: Stealthy Memory Corruption in LLM-Powered Recommender Agents

2025-03-31 · Shiyi Yang, Zhibo Hu, Xinshu Li, Chen Wang 외

Large language model (LLM)-powered agents are increasingly used in recommender systems (RSs) to achieve personalized behavior modeling, where the memory mechanism plays a pivotal role in enabling the agents to autonomous…

Collaborative FilteringLarge Language ModelRecommendation Systems

Sponge Tool Attack: Stealthy Denial-of-Efficiency against Tool-Augmented Agentic Reasoning

2026-01-24 · Qi Li, Xinchao Wang arxiv

Enabling large language models (LLMs) to solve complex reasoning tasks is a key step toward artificial general intelligence. Recent work augments LLMs with external tools to enable agentic reasoning, achieving high utili…

SABER: Symbolic Regression-based Angle of Arrival and Beam Pattern Estimator

2025-10-30 · Shih-Kai Chou, Mengran Zhao, Cheng-Nan Hu, Kuang-Chung Chou 외 arxiv

Accurate Angle-of-arrival (AoA) estimation is essential for next-generation wireless communication systems to enable reliable beamforming, high-precision localization, and integrated sensing. Unfortunately, classical hig…

Silent Killer: A Stealthy, Clean-Label, Black-Box Backdoor Attack

2023-01-05 · Tzvi Lederer, Gallil Maimon, Lior Rokach

Backdoor poisoning attacks pose a well-known risk to neural networks. However, most studies have focused on lenient threat models. We introduce Silent Killer, a novel attack that operates in clean-label, black-box settin…

Backdoor AttackData Poisoning