paper-with-me

Papers

Safe Reinforcement Learning in Tensor Reproducing Kernel Hilbert Space

2023-12-01 · Xiaoyuan Cheng, Boli Chen, Liz Varga, Yukun Hu

This paper delves into the problem of safe reinforcement learning (RL) in a partially observable environment with the aim of achieving safe-reachability objectives. In traditional partially observable Markov decision processes (POMDP), ensuring safety typically involves estimating the belief in latent states. However, accurately estimating an optimal Bayesian filter in POMDP to infer latent states from observations in a continuous state space poses a significant challenge, largely due to the intractable likelihood. To tackle this issue, we propose a stochastic model-based approach that guarantees RL safety almost surely in the face of unknown system dynamics and partial observation environments. We leveraged the Predictive State Representation (PSR) and Reproducing Kernel Hilbert Space (RKHS) to represent future multi-step observations analytically, and the results in this context are provable. Furthermore, we derived essential operators from the kernel Bayes' rule, enabling the recursive estimation of future observations using various operators. Under the assumption of \textit{undercompleness}, a polynomial sample complexity is established for the RL algorithm for the infinite size of observation and action spaces, ensuring an $\epsilon-$suboptimal safe policy guarantee.

📄 PDF Abstract BibTeX arXiv:2312.00727

Code (0)

등록된 구현이 없습니다.

Tasks

reinforcement-learningReinforcement LearningReinforcement Learning (RL)Safe Reinforcement Learning

Similar Papers 제목 키워드 기반

The Matrix Hilbert Space and Its Application to Matrix Learning

2017-06-25 · Yunfei Ye

Theoretical studies have proven that the Hilbert space has remarkable performance in many fields of applications. Frames in tensor product of Hilbert spaces were introduced to generalize the inner product to high-order t…

Tensor Decomposition

Irregular Sampling of High-Dimensional Functions in Reproducing Kernel Hilbert Spaces

2025-04-18 · Armin Iske, Lennart Ohlsen

We develop sampling formulas for high-dimensional functions in reproducing kernel Hilbert spaces, where we rely on irregular samples that are taken at determining sequences of data points. We place particular emphasis on…

Learning Tensors in Reproducing Kernel Hilbert Spaces with Multilinear Spectral Penalties

2013-10-18 · Marco Signoretto, Lieven De Lathauwer, Johan A. K. Suykens

We present a general framework to learn functions in tensor product reproducing kernel Hilbert spaces (TP-RKHSs). The methodology is based on a novel representer theorem suitable for existing as well as new spectral pena…

Transfer Learning

Safe exploration in reproducing kernel Hilbert spaces

2025-03-13 · Abdullah Tokmak, Kiran G. Krishnan, Thomas B. Schön, Dominik Baumann

Popular safe Bayesian optimization (BO) algorithms learn control policies for safety-critical systems in unknown environments. However, most algorithms make a smoothness assumption, which is encoded by a known bounded no…

Bayesian OptimizationSafe Exploration

Representation of Reinforcement Learning Policies in Reproducing Kernel Hilbert Spaces

2020-02-07 · Bogdan Mazoure, Thang Doan, Tianyu Li, Vladimir Makarenkov 외

We propose a general framework for policy representation for reinforcement learning tasks. This framework involves finding a low-dimensional embedding of the policy on a reproducing kernel Hilbert space (RKHS). The usage…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)