paper-with-me

Papers

Balancing detectability and performance of attacks on the control channel of Markov Decision Processes

2021-09-15 · Alessio Russo, Alexandre Proutiere

We investigate the problem of designing optimal stealthy poisoning attacks on the control channel of Markov decision processes (MDPs). This research is motivated by the recent interest of the research community for adversarial and poisoning attacks applied to MDPs, and reinforcement learning (RL) methods. The policies resulting from these methods have been shown to be vulnerable to attacks perturbing the observations of the decision-maker. In such an attack, drawing inspiration from adversarial examples used in supervised learning, the amplitude of the adversarial perturbation is limited according to some norm, with the hope that this constraint will make the attack imperceptible. However, such constraints do not grant any level of undetectability and do not take into account the dynamic nature of the underlying Markov process. In this paper, we propose a new attack formulation, based on information-theoretical quantities, that considers the objective of minimizing the detectability of the attack as well as the performance of the controlled process. We analyze the trade-off between the efficiency of the attack and its detectability. We conclude with examples and numerical simulations illustrating this trade-off.

📄 PDF Abstract BibTeX arXiv:2109.07171

Code (1)

rssalessio/optimal-attack-control-channel-mdp 공식 구현

Tasks

Reinforcement Learning (RL)

Similar Papers 제목 키워드 기반

Cyber-Attack Detection in Discrete Nonlinear Multi-Agent Systems Using Neural Networks

2021-02-01 · Amirreza Mousavi, Kiarash Aryankia, Rastko R. Selmic

This paper proposes a distributed cyber-attack detection method in communication channels for a class of discrete, nonlinear, heterogeneous, multi-agent systems that are controlled by our proposed formation-based control…

Cyber Attack Detection

Illusory Attacks: Information-Theoretic Detectability Matters in Adversarial Attacks

2022-07-20 · Tim Franzmeyer, Stephen Mcaleer, João F. Henriques, Jakob N. Foerster 외

Autonomous agents deployed in the real world need to be robust against adversarial attacks on sensory inputs. Robustifying agent policies requires anticipating the strongest attacks possible. We demonstrate that existing…

Adversarial AttackAdversarial Robustness

Robust LLM Watermarking with Minimal Semantic Distortion for IP Protection

2026-05-22 · Kieu Dang, Phung Lai, NhatHai Phan, Yelong Shen 외 arxiv

Proprietary large language models (LLMs) face risks of intellectual property (IP) violation, as adversaries can replicate an LLM by collecting input-output pairs to train a surrogate model, causing financial setbacks. Wa…

Improving Generalizability and Undetectability for Targeted Adversarial Attacks on Multimodal Pre-trained Models

2025-09-24 · Zhifang Zhang, Jiahan Zhang, Shengjie Zhou, Qi Wei 외 arxiv

Multimodal pre-trained models (e.g., ImageBind), which align distinct data modalities into a shared embedding space, have shown remarkable success across downstream tasks. However, their increasing adoption raises seriou…

Anomaly Detection

Enhancing sensor attack detection in supervisory control systems modeled by probabilistic automata

2025-02-23 · Parastou Fahim, Samuel Oliveira, Rômulo Meira-Góes

Sensor attacks compromise the reliability of cyber-physical systems (CPSs) by altering sensor outputs with the objective of leading the system to unsafe system states. This paper studies a probabilistic intrusion detecti…

Intrusion Detection