paper-with-me

홈 › Papers

PoolFlip: A Multi-Agent Reinforcement Learning Security Environment for Cyber Defense

2025-08-27 · Xavier Cadet, Simona Boboila, Sie Hendrata Dharmawan, Alina Oprea, Peter Chin arxiv

Cyber defense requires automating defensive decision-making under stealthy, deceptive, and continuously evolving adversarial strategies. The FlipIt game provides a foundational framework for modeling interactions between a defender and an advanced adversary that compromises a system without being immediately detected. In FlipIt, the attacker and defender compete to control a shared resource by performing a Flip action and paying a cost. However, the existing FlipIt frameworks rely on a small number of heuristics or specialized learning techniques, which can lead to brittleness and the inability to adapt to new attacks. To address these limitations, we introduce PoolFlip, a multi-agent gym environment that extends the FlipIt game to allow efficient learning for attackers and defenders. Furthermore, we propose Flip-PSRO, a multi-agent reinforcement learning (MARL) approach that leverages population-based training to train defender agents equipped to generalize against a range of unknown, potentially adaptive opponents. Our empirical results suggest that Flip-PSRO defenders are $2\times$ more effective than baselines to generalize to a heuristic attack not exposed in training. In addition, our newly designed ownership-based utility functions ensure that Flip-PSRO defenders maintain a high level of control while optimizing performance.

📄 PDF Abstract BibTeX arXiv:2508.19488

Code (0)

등록된 구현이 없습니다.

Tasks

Multi-agent Reinforcement Learning

Similar Papers 제목 키워드 기반

Learning Communication Between Heterogeneous Agents in Multi-Agent Reinforcement Learning for Autonomous Cyber Defence

2026-03-17 · Alex Popa, Adrian Taylor, Ranwa Al Mallah arxiv

Reinforcement learning techniques are being explored as solutions to the threat of cyber attacks on enterprise networks. Recent research in the field of AI in cyber security has investigated the ability of homogeneous mu…

Multi-agent Reinforcement Learning

Out of the Cage: How Stochastic Parrots Win in Cyber Security Environments

2023-08-23 · Maria Rigaki, Ondřej Lukáš, Carlos A. Catania, Sebastian Garcia

Large Language Models (LLMs) have gained widespread popularity across diverse domains involving text generation, summarization, and various natural language processing tasks. Despite their inherent limitations, LLM-based…

CyberBattleSimCyberBattleSim (RL) chain scenarioDecision MakingNetSecGame+6

Reinforcement Learning for Hardware Security: Opportunities, Developments, and Challenges

2022-08-29 · Satwik Patnaik, Vasudev Gohil, Hao Guo, Jeyavijayan 외

Reinforcement learning (RL) is a machine learning paradigm where an autonomous agent learns to make an optimal sequence of decisions by interacting with the underlying environment. The promise demonstrated by RL-guided w…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Policy Teaching in Reinforcement Learning via Environment Poisoning Attacks

2020-11-21 · Amin Rakhsha, Goran Radanovic, Rati Devidze, Xiaojin Zhu 외

We study a security threat to reinforcement learning where an attacker poisons the learning environment to force the agent into executing a target policy chosen by the attacker. As a victim, we consider RL agents whose o…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

PolicyCleanse: Backdoor Detection and Mitigation in Reinforcement Learning

2022-02-08 · Junfeng Guo, Ang Li, Cong Liu

While real-world applications of reinforcement learning are becoming popular, the security and robustness of RL systems are worthy of more attention and exploration. In particular, recent works have revealed that, in a m…

Machine Unlearningreinforcement-learningReinforcement LearningReinforcement Learning (RL)