paper-with-me

홈 › Papers

Learning Finite State Representations of Recurrent Policy Networks

2018-11-29 · ICLR 2019 5 · Anurag Koul, Sam Greydanus, Alan Fern

Recurrent neural networks (RNNs) are an effective representation of control policies for a wide range of reinforcement and imitation learning problems. RNN policies, however, are particularly difficult to explain, understand, and analyze due to their use of continuous-valued memory vectors and observation features. In this paper, we introduce a new technique, Quantized Bottleneck Insertion, to learn finite representations of these vectors and features. The result is a quantized representation of the RNN that can be analyzed to improve our understanding of memory use and general behavior. We present results of this approach on synthetic environments and six Atari games. The resulting finite representations are surprisingly small in some cases, using as few as 3 discrete memory states and 10 observations for a perfect Pong policy. We also show that these finite policy representations lead to improved interpretability.

📄 PDF Abstract BibTeX arXiv:1811.12530

Code (0)

등록된 구현이 없습니다.

Tasks

Atari GamesImitation Learning

Similar Papers 제목 키워드 기반

Re-understanding Finite-State Representations of Recurrent Policy Networks

2020-06-06 · Mohamad H. Danesh, Anurag Koul, Alan Fern, Saeed Khorram

We introduce an approach for understanding control policies represented as recurrent neural networks. Recent work has approached this problem by transforming such recurrent policy networks into finite-state machines (FSM…

Atari Games

Recurrent Predictive State Policy Networks

2018-03-05 · ICML 2018 7 · Ahmed Hefny, Zita Marinho, Wen Sun, Siddhartha Srinivasa 외

We introduce Recurrent Predictive State Policy (RPSP) networks, a recurrent architecture that brings insights from predictive state representations to reinforcement learning in partially observable environments. Predicti…

OpenAI GymReinforcement Learning

Recurrent Natural Policy Gradient for POMDPs

2024-05-28 · Semih Cayci, Atilla Eryilmaz

In this paper, we study a natural policy gradient method based on recurrent neural networks (RNNs) for partially-observable Markov decision processes, whereby RNNs are used for policy parameterization and policy evaluati…

Recurrent Model Predictive Control

2021-02-23 · Zhengyu Liu, Jingliang Duan, Wenxuan Wang, Shengbo Eben Li 외

This paper proposes an off-line algorithm, called Recurrent Model Predictive Control (RMPC), to solve general nonlinear finite-horizon optimal control problems. Unlike traditional Model Predictive Control (MPC) algorithm…

modelModel Predictive Control

Recurrent Model Predictive Control: Learning an Explicit Recurrent Controller for Nonlinear Systems

2021-02-20 · Zhengyu Liu, Jingliang Duan, Wenxuan Wang, Shengbo Eben Li 외

This paper proposes an offline control algorithm, called Recurrent Model Predictive Control (RMPC), to solve large-scale nonlinear finite-horizon optimal control problems. It can be regarded as an explicit solver of trad…

Model Predictive Control