paper-with-me

Papers

BACKDOORL: Backdoor Attack against Competitive Reinforcement Learning

2021-05-02 · Lun Wang, Zaynah Javed, Xian Wu, Wenbo Guo, Xinyu Xing, Dawn Song

Recent research has confirmed the feasibility of backdoor attacks in deep reinforcement learning (RL) systems. However, the existing attacks require the ability to arbitrarily modify an agent's observation, constraining the application scope to simple RL systems such as Atari games. In this paper, we migrate backdoor attacks to more complex RL systems involving multiple agents and explore the possibility of triggering the backdoor without directly manipulating the agent's observation. As a proof of concept, we demonstrate that an adversary agent can trigger the backdoor of the victim agent with its own action in two-player competitive RL systems. We prototype and evaluate BACKDOORL in four competitive environments. The results show that when the backdoor is activated, the winning rate of the victim drops by 17% to 37% compared to when not activated.

📄 PDF Abstract BibTeX arXiv:2105.00579

Code (0)

등록된 구현이 없습니다.

Tasks

Atari GamesBackdoor AttackDeep Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

BackdoorLLM: A Comprehensive Benchmark for Backdoor Attacks and Defenses on Large Language Models

2024-08-23 · Yige Li, Hanxun Huang, Yunhan Zhao, Xingjun Ma 외

Generative large language models (LLMs) have achieved state-of-the-art results on a wide range of tasks, yet they remain susceptible to backdoor attacks: carefully crafted triggers in the input can manipulate the model t…

Data Poisoningtext-classificationText ClassificationText Generation

AutoBackdoor: Automating Backdoor Attacks via LLM Agents

2025-11-20 · Yige Li, Zhe Li, Wei Zhao, Nay Myat Min 외 arxiv

Backdoor attacks pose a serious threat to the secure deployment of large language models (LLMs), enabling adversaries to implant hidden behaviors triggered by specific inputs. However, existing methods often rely on manu…

Backdoor4Good: Benchmarking Beneficial Uses of Backdoors in LLMs

2026-03-08 · Yige Li, Wei Zhao, Zhe Li, Nay Myat Min 외 arxiv

Backdoor mechanisms have traditionally been studied as security threats that compromise the integrity of machine learning models. However, the same mechanism -- the conditional activation of specific behaviors through in…

A Spatiotemporal Stealthy Backdoor Attack against Cooperative Multi-Agent Deep Reinforcement Learning

2024-09-12 · Yinbo Yu, Saihao Yan, Jiajia Liu

Recent studies have shown that cooperative multi-agent deep reinforcement learning (c-MADRL) is under the threat of backdoor attacks. Once a backdoor trigger is observed, it will perform abnormal actions leading to failu…

Backdoor AttackDeep Reinforcement LearningSMACSMAC+

BLAST: A Stealthy Backdoor Leverage Attack against Cooperative Multi-Agent Deep Reinforcement Learning based Systems

2025-01-03 · Yinbo Yu, Saihao Yan, Xueyu Yin, Jing Fang 외

Recent studies have shown that cooperative multi-agent deep reinforcement learning (c-MADRL) is under the threat of backdoor attacks. Once a backdoor trigger is observed, it will perform malicious actions leading to fail…

Deep Reinforcement LearningSMACSMAC+