paper-with-me

홈 › Papers

MAGICS: Adversarial RL with Minimax Actors Guided by Implicit Critic Stackelberg for Convergent Neural Synthesis of Robot Safety

2024-09-20 · Justin Wang, Haimin Hu, Duy Phuong Nguyen, Jaime Fernández Fisac

While robust optimal control theory provides a rigorous framework to compute robot control policies that are provably safe, it struggles to scale to high-dimensional problems, leading to increased use of deep learning for tractable synthesis of robot safety. Unfortunately, existing neural safety synthesis methods often lack convergence guarantees and solution interpretability. In this paper, we present Minimax Actors Guided by Implicit Critic Stackelberg (MAGICS), a novel adversarial reinforcement learning (RL) algorithm that guarantees local convergence to a minimax equilibrium solution. We then build on this approach to provide local convergence guarantees for a general deep RL-based robot safety synthesis algorithm. Through both simulation studies on OpenAI Gym environments and hardware experiments with a 36-dimensional quadruped robot, we show that MAGICS can yield robust control policies outperforming the state-of-the-art neural safety synthesis methods.

📄 PDF Abstract BibTeX arXiv:2409.13867

Code (0)

등록된 구현이 없습니다.

Tasks

OpenAI GymReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

GDA-AM: On the effectiveness of solving minimax optimization via Anderson Acceleration

2021-10-06 · Huan He, Shifan Zhao, Yuanzhe Xi, Joyce C Ho 외

Many modern machine learning algorithms such as generative adversarial networks (GANs) and adversarial training can be formulated as minimax optimization. Gradient descent ascent (GDA) is the most commonly used algorithm…

Nonparametric Density Estimation under Adversarial Losses

2018-05-22 · NeurIPS 2018 12 · Shashank Singh, Ananya Uppal, Boyue Li, Chun-Liang Li 외

We study minimax convergence rates of nonparametric density estimation under a large class of loss functions called "adversarial losses", which, besides classical $\mathcal{L}^p$ losses, includes maximum mean discrepancy…

Density Estimation

GDA-AM: ON THE EFFECTIVENESS OF SOLVING MIN-IMAX OPTIMIZATION VIA ANDERSON MIXING

2021-09-29 · ICLR 2022 4 · Huan He, Shifan Zhao, Yuanzhe Xi, Joyce Ho 외

Many modern machine learning algorithms such as generative adversarial networks (GANs) and adversarial training can be formulated as minimax optimization.Gradient descent ascent (GDA) is the most commonly used algorithm …

Semi-Implicit Hybrid Gradient Methods with Application to Adversarial Robustness

2022-02-21 · Beomsu Kim, Junghoon Seo

Adversarial examples, crafted by adding imperceptible perturbations to natural inputs, can easily fool deep neural networks (DNNs). One of the most successful methods for training adversarially robust DNNs is solving a n…

Adversarial Robustness

Adversarially Robust Learning for Security-Constrained Optimal Power Flow

2021-11-12 · NeurIPS 2021 12 · Priya L. Donti, Aayushya Agarwal, Neeraj Vijay Bedmutha, Larry Pileggi 외

In recent years, the ML community has seen surges of interest in both adversarially robust learning and implicit layers, but connections between these two areas have seldom been explored. In this work, we combine innovat…