paper-with-me

Papers

Continuous-time Value Function Approximation in Reproducing Kernel Hilbert Spaces

2018-06-08 · NeurIPS 2018 12 · Motoya Ohnishi, Masahiro Yukawa, Mikael Johansson, Masashi Sugiyama

Motivated by the success of reinforcement learning (RL) for discrete-time tasks such as AlphaGo and Atari games, there has been a recent surge of interest in using RL for continuous-time control of physical systems (cf. many challenging tasks in OpenAI Gym and DeepMind Control Suite). Since discretization of time is susceptible to error, it is methodologically more desirable to handle the system dynamics directly in continuous time. However, very few techniques exist for continuous-time RL and they lack flexibility in value function approximation. In this paper, we propose a novel framework for model-based continuous-time value function approximation in reproducing kernel Hilbert spaces. The resulting framework is so flexible that it can accommodate any kind of kernel-based approach, such as Gaussian processes and kernel adaptive filters, and it allows us to handle uncertainties and nonstationarity without prior knowledge about the environment or what basis functions to employ. We demonstrate the validity of the presented framework through experiments.

📄 PDF Abstract BibTeX arXiv:1806.02985

Code (0)

등록된 구현이 없습니다.

Tasks

Atari GamesGaussian ProcessesOpenAI GymReinforcement LearningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

Strictly Decentralized Adaptive Estimation of External Fields using Reproducing Kernels

2021-03-23 · Jia Guo, Michael E. Kepler, Sai Tej Paruchuri, Haoran Wang 외

This paper describes an adaptive method in continuous time for the estimation of external fields by a team of $N$ agents. The agents $i$ each explore subdomains $\Omega^i$ of a bounded subset of interest $\Omega\subset X…

Unity

Conditioning of Banach Space Valued Gaussian Random Variables: An Approximation Approach Based on Martingales

2024-04-04 · Ingo Steinwart

We investigate the conditional distributions of two Banach space valued, jointly Gaussian random variables. In particular, we show that these conditional distributions are again Gaussian and that their means and covarian…

Gaussian Processes

On the Convergence of Irregular Sampling in Reproducing Kernel Hilbert Spaces

2025-04-18 · Armin Iske

We analyse the convergence of sampling algorithms for functions in reproducing kernel Hilbert spaces (RKHS). To this end, we discuss approximation properties of kernel regression under minimalistic assumptions on both th…

regression

Global universal approximation of functional input maps on weighted spaces

2023-06-05 · Christa Cuchiero, Philipp Schmocker, Josef Teichmann

We introduce so-called functional input neural networks defined on a possibly infinite dimensional weighted space with values also in a possibly infinite dimensional output space. To this end, we use an additive family t…

Gaussian ProcessesregressionUncertainty Quantification

Rates of Convergence in Certain Native Spaces of Approximations used in Reinforcement Learning

2023-09-14 · Ali Bouland, Shengyuan Niu, Sai Tej Paruchuri, Andrew Kurdila 외

This paper studies convergence rates for some value function approximations that arise in a collection of reproducing kernel Hilbert spaces (RKHS) $H(\Omega)$. By casting an optimal control problem in a specific class of…