paper-with-me

Papers

BadRobot: Jailbreaking Embodied LLMs in the Physical World

2024-07-16 · Hangtao Zhang, Chenyu Zhu, Xianlong Wang, Ziqi Zhou, Changgan Yin, Minghui Li, Lulu Xue, Yichen Wang, Shengshan Hu, Aishan Liu, Peijin Guo, Leo Yu Zhang

Embodied AI represents systems where AI is integrated into physical entities. Large Language Model (LLM), which exhibits powerful language understanding abilities, has been extensively employed in embodied AI by facilitating sophisticated task planning. However, a critical safety issue remains overlooked: could these embodied LLMs perpetrate harmful behaviors? In response, we introduce BadRobot, a novel attack paradigm aiming to make embodied LLMs violate safety and ethical constraints through typical voice-based user-system interactions. Specifically, three vulnerabilities are exploited to achieve this type of attack: (i) manipulation of LLMs within robotic systems, (ii) misalignment between linguistic outputs and physical actions, and (iii) unintentional hazardous behaviors caused by world knowledge's flaws. Furthermore, we construct a benchmark of various malicious physical action queries to evaluate BadRobot's attack performance. Based on this benchmark, extensive experiments against existing prominent embodied LLM frameworks (e.g., Voxposer, Code as Policies, and ProgPrompt) demonstrate the effectiveness of our BadRobot.

📄 PDF Abstract BibTeX arXiv:2407.20242

Code (0)

등록된 구현이 없습니다.

Tasks

Language ModelingLanguage ModellingLarge Language ModelTask Planning

Similar Papers 제목 키워드 기반

Jailbreaking Embodied LLMs via Action-level Manipulation

2026-03-02 · Xinyu Huang, Qiang Yang, Leming Shen, Zijing Ma 외 arxiv

Embodied Large Language Models (LLMs) enable AI agents to interact with the physical world through natural language instructions and actions. However, beyond the language-level risks inherent to LLMs themselves, embodied…

Concept-Based Dictionary Learning for Inference-Time Safety in Vision Language Action Models

2026-02-02 · Siqi Wen, Shu Yang, Shaopeng Fu, Jingfeng Zhang 외 arxiv

Vision Language Action (VLA) models close the perception action loop by translating multimodal instructions into executable behaviors, but this very capability magnifies safety risks: jailbreaks that merely yield toxic t…

Embodied AI: From LLMs to World Models

2025-09-24 · Tongtong Feng, Xin Wang, Yu-Gang Jiang, Wenwu Zhu arxiv

Embodied Artificial Intelligence (AI) is an intelligent system paradigm for achieving Artificial General Intelligence (AGI), serving as the cornerstone for various applications and driving the evolution from cyberspace t…

Beyond APIs: Probing the Limits of MLLMs in Physical Tool Use

2026-06-09 · Zhixin Ma, Yutong Zhou, Yongqi Li, Chong-Wah Ngo 외 arxiv

Multimodal Large Language Models (MLLMs) excel at utilizing digital APIs and increasingly serve as the "brain" of embodied AI, instructing robots to interact with the physical world. In such embodied settings, a central …

Trust in LLM-controlled Robotics: a Survey of Security Threats, Defenses and Challenges

2025-12-17 · Xinyu Huang, Shyam Karthick V B, Taozhao Chen, Mitch Bryson 외 arxiv

The integration of Large Language Models (LLMs) into robotics has revolutionized their ability to interpret complex human commands and execute sophisticated tasks. However, such paradigm shift introduces critical securit…