paper-with-me

Papers

Theoretical understanding of adversarial reinforcement learning via mean-field optimal control

2021-09-29 · ZiMing Wang, Fengxiang He, Tao Cui, DaCheng Tao

Adversarial reinforcement learning has been shown promising in solving games in adversarial environments, while the theoretical understanding is still premature. This paper theoretically analyses the convergence and generalization of adversarial reinforcement learning under the mean-field optimal control framework. A new mean-field Pontryagin's maximum principle is proposed for reinforcement learning with implicit terminal constraints. Applying Hamilton-Jacobi-Issacs equation and mean-field two-sided extremism principle (TSEP), adversarial reinforcement learning is modeled as a mean-field quantitative differential game between two constrained dynamical systems. These results provide the necessary conditions for the convergence of the global solution to the mean-field TSEP. The global solution is also unique when the terminal time is sufficiently small. Moreover, two generalization bounds are delivered via Hoeffding's inequality and algorithmic stability. Both bounds do not explicitly depend on the dimensions, norms, or other capacity measures of the parameter, which are usually prohibitively large in deep learning. The bounds help characterize how the algorithm randomness facilitates the generalization of adversarial reinforcement learning. Moreover, the techniques may be helpful in modeling other adversarial learning algorithms.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Generalization Boundsreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

Adversarial Training from Mean Field Perspective

2025-05-20 · NeurIPS 2023 11 · Soichiro Kumano, Hiroshi Kera, Toshihiko Yamasaki

Although adversarial training is known to be effective against adversarial examples, training dynamics are not well understood. In this study, we present the first theoretical analysis of adversarial training in random d…

RoMFAC: A robust mean-field actor-critic reinforcement learning against adversarial perturbations on states

2022-05-15 · Ziyuan Zhou, Guanjun Liu

Multi-agent deep reinforcement learning makes optimal decisions dependent on system states observed by agents, but any uncertainty on the observations may mislead agents to take wrong actions. The Mean-Field Actor-Critic…

Deep Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Discrete-Time Mean Field Control with Environment States

2021-04-30 · Kai Cui, Anam Tahir, Mark Sinzger, Heinz Koeppl

Multi-agent reinforcement learning methods have shown remarkable potential in solving complex multi-agent problems but mostly lack theoretical guarantees. Recently, mean field control and mean field games have been estab…

Deep Reinforcement LearningMulti-agent Reinforcement Learningreinforcement-learningReinforcement Learning+2

Adversarial Inverse Reinforcement Learning for Mean Field Games

2021-04-29 · Yang Chen, Libo Zhang, Jiamou Liu, Michael Witbrock

Mean field games (MFGs) provide a mathematically tractable framework for modelling large-scale multi-agent systems by leveraging mean field theory to simplify interactions among agents. It enables applying inverse reinfo…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Error Bounds of Imitating Policies and Environments

2020-10-22 · NeurIPS 2020 12 · Tian Xu, Ziniu Li, Yang Yu

Imitation learning trains a policy by mimicking expert demonstrations. Various imitation methods were proposed and empirically evaluated, meanwhile, their theoretical understanding needs further studies. In this paper, w…

Imitation LearningModel-based Reinforcement Learningreinforcement-learningReinforcement Learning+1