paper-with-me

홈 › Papers

Online Robust Policy Learning in the Presence of Unknown Adversaries

2018-07-16 · NeurIPS 2018 12 · Aaron J. Havens, Zhanhong Jiang, Soumik Sarkar

The growing prospect of deep reinforcement learning (DRL) being used in cyber-physical systems has raised concerns around safety and robustness of autonomous agents. Recent work on generating adversarial attacks have shown that it is computationally feasible for a bad actor to fool a DRL policy into behaving sub optimally. Although certain adversarial attacks with specific attack models have been addressed, most studies are only interested in off-line optimization in the data space (e.g., example fitting, distillation). This paper introduces a Meta-Learned Advantage Hierarchy (MLAH) framework that is attack model-agnostic and more suited to reinforcement learning, via handling the attacks in the decision space (as opposed to data space) and directly mitigating learned bias introduced by the adversary. In MLAH, we learn separate sub-policies (nominal and adversarial) in an online manner, as guided by a supervisory master agent that detects the presence of the adversary by leveraging the advantage function for the sub-policies. We demonstrate that the proposed algorithm enables policy learning with significantly lower bias as compared to the state-of-the-art policy learning approaches even in the presence of heavy state information attacks. We present algorithm analysis and simulation results using popular OpenAI Gym environments.

📄 PDF Abstract BibTeX arXiv:1807.06064

Code (0)

등록된 구현이 없습니다.

Tasks

Deep Reinforcement LearningOpenAI Gymreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

Integrating Planning, Execution and Monitoring in the presence of Open World Novelties: Case Study of an Open World Monopoly Solver

2021-07-09 · Sriram Gopalakrishnan, Utkarsh Soni, Tung Thai, Panagiotis Lymperopoulos 외

The game of monopoly is an adversarial multi-agent domain where there is no fixed goal other than to be the last player solvent, There are useful subgoals like monopolizing sets of properties, and developing them. There …

Physical-Layer Security via Distributed Beamforming in the Presence of Adversaries with Unknown Locations

2021-02-28 · Yagiz Savas, Abolfazl Hashemi, Abraham P. Vinod, Brian M. Sadler 외

We study the problem of securely communicating a sequence of information bits with a client in the presence of multiple adversaries at unknown locations in the environment. We assume that the client and the adversaries a…

Online Learning with Adversaries: A Differential-Inclusion Analysis

2023-04-04 · Swetha Ganesh, Alexandre Reiffers-Masson, Gugan Thoppe

We introduce an observation-matrix-based framework for fully asynchronous online Federated Learning (FL) with adversaries. In this work, we demonstrate its effectiveness in estimating the mean of a random vector. Our mai…

Federated Learning

Extend Adversarial Policy Against Neural Machine Translation via Unknown Token

2025-01-21 · Wei Zou, ShuJian Huang, Jiajun Chen

Generating adversarial examples contributes to mainstream neural machine translation~(NMT) robustness. However, popular adversarial policies are apt for fixed tokenization, hindering its efficacy for common character per…

Machine TranslationNMTReinforcement Learning (RL)Translation

Data-Guided Regulator for Adaptive Nonlinear Control

2023-11-20 · Niyousha Rahimi, Mehran Mesbahi

This paper addresses the problem of designing a data-driven feedback controller for complex nonlinear dynamical systems in the presence of time-varying disturbances with unknown dynamics. Such disturbances are modeled as…

Time Series