paper-with-me

Papers

Using Echo State Networks to Approximate Value Functions for Control

2021-02-11 · Allen G. Hart, Kevin R. Olding, A. M. G. Cox, Olga Isupova, J. H. P. Dawes

An Echo State Network (ESN) is a type of single-layer recurrent neural network with randomly-chosen internal weights and a trainable output layer. We prove under mild conditions that a sufficiently large Echo State Network can approximate the value function of a broad class of stochastic and deterministic control problems. Such control problems are generally non-Markovian. We describe how the ESN can form the basis for novel and computationally efficient reinforcement learning algorithms in a non-Markovian framework. We demonstrate this theory with two examples. In the first, we use an ESN to solve a deterministic, partially observed, control problem which is a simple game we call `Bee World'. In the second example, we consider a stochastic control problem inspired by a market making problem in mathematical finance. In both cases we can compare the dynamics of the algorithms with analytic solutions to show that even after only a single reinforcement policy iteration the algorithms arrive at a good policy.

📄 PDF Abstract BibTeX arXiv:2102.06258

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Echo State Condition at the Critical Point

2014-11-25 · Norbert Michael Mayer

Recurrent networks with transfer functions that fulfill the Lipschitz continuity with K=1 may be echo state networks if certain limitations on the recurrent connectivity are applied. It has been shown that it is sufficie…

Universality and approximation bounds for echo state networks with random weights

2022-06-12 · Zhen Li, Yunfei Yang

We study the uniform approximation of echo state networks with randomly generated internal weights. These models, in which only the readout weights are optimized during training, have made empirical success in learning d…

Domain Shift in Echocardiography: Interpretable Quantification and Prediction of Cross-Dataset Left Ventricular Segmentation

2026-07-22 · Soroush Elyasi, Nasim Dadashi Serej, Julie Wall, Massoud Zolgharni arxiv

Cross-dataset generalisation remains a major barrier to clinical deployment of echocardiographic left ventricular segmentation, yet the sources of this shift are rarely disentangled. We examined whether transfer degradat…

Difference of Convex Functions Programming for Reinforcement Learning

2014-12-01 · NeurIPS 2014 12 · Bilal Piot, Matthieu Geist, Olivier Pietquin

Large Markov Decision Processes (MDPs) are usually solved using Approximate Dynamic Programming (ADP) methods such as Approximate Value Iteration (AVI) or Approximate Policy Iteration (API). The main contribution of this…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Efficient Acoustic Echo Suppression with Condition-Aware Training

2023-07-28 · Ernst Seidel, Pejman Mowlaee, Tim Fingscheidt

The topic of deep acoustic echo control (DAEC) has seen many approaches with various model topologies in recent years. Convolutional recurrent networks (CRNs), consisting of a convolutional encoder and decoder encompassi…

Decoder