paper-with-me

홈 › Papers

Perturbational Complexity by Distribution Mismatch: A Systematic Analysis of Reinforcement Learning in Reproducing Kernel Hilbert Space

2021-11-05 · Jihao Long, Jiequn Han

Most existing theoretical analysis of reinforcement learning (RL) is limited to the tabular setting or linear models due to the difficulty in dealing with function approximation in high dimensional space with an uncertain environment. This work offers a fresh perspective into this challenge by analyzing RL in a general reproducing kernel Hilbert space (RKHS). We consider a family of Markov decision processes $\mathcal{M}$ of which the reward functions lie in the unit ball of an RKHS and transition probabilities lie in a given arbitrary set. We define a quantity called perturbational complexity by distribution mismatch $\Delta_{\mathcal{M}}(\epsilon)$ to characterize the complexity of the admissible state-action distribution space in response to a perturbation in the RKHS with scale $\epsilon$. We show that $\Delta_{\mathcal{M}}(\epsilon)$ gives both the lower bound of the error of all possible algorithms and the upper bound of two specific algorithms (fitted reward and fitted Q-iteration) for the RL problem. Hence, the decay of $\Delta_\mathcal{M}(\epsilon)$ with respect to $\epsilon$ measures the difficulty of the RL problem on $\mathcal{M}$. We further provide some concrete examples and discuss whether $\Delta_{\mathcal{M}}(\epsilon)$ decays fast or not in these examples. As a byproduct, we show that when the reward functions lie in a high dimensional RKHS, even if the transition probability is known and the action space is finite, it is still possible for RL problems to suffer from the curse of dimensionality.

📄 PDF Abstract BibTeX arXiv:2111.03469

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning (RL)

Similar Papers 제목 키워드 기반

Reinforcement Learning with Function Approximation: From Linear to Nonlinear

2023-02-20 · Jihao Long, Jiequn Han

Function approximation has been an indispensable component in modern reinforcement learning algorithms designed to tackle problems with large state spaces in high dimensions. This paper reviews recent results on error an…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Intervention-Aware Multiscale Representation Learning from Imaging Phenomics and Perturbation Transcriptomics

2026-04-19 · Jiayuan Chen, Ruoqi Liu, Zishan Gu, Ping Zhang arxiv

Microscopy-based phenotypic profiling is scalable for drug discovery but lacks the mechanistic depth of transcriptomics, which remains costly and scarce. Existing multimodal approaches either use images to support other …

Representation LearningDrug Discovery

Off-Policy Policy Gradient Algorithms by Constraining the State Distribution Shift

2019-11-16 · Riashat Islam, Komal K. Teru, Deepak Sharma, Joelle Pineau

Off-policy deep reinforcement learning (RL) algorithms are incapable of learning solely from batch offline data without online interactions with the environment, due to the phenomenon known as \textit{extrapolation error…

continuous-controlContinuous ControlDeep Reinforcement LearningReinforcement Learning+1

Deep Model-Based Architectures for Inverse Problems under Mismatched Priors

2022-07-26 · Shirin Shoushtari, Jiaming Liu, Yuyang Hu, Ulugbek S. Kamilov

There is a growing interest in deep model-based architectures (DMBAs) for solving imaging inverse problems by combining physical measurement models and learned image priors specified using convolutional neural nets (CNNs…

MixMOOD: A systematic approach to class distribution mismatch in semi-supervised learning using deep dataset dissimilarity measures

2020-06-14 · Saul Calderon-Ramirez, Luis Oala, Jordina Torrents-Barrena, Shengxiang Yang 외

In this work, we propose MixMOOD - a systematic approach to mitigate effect of class distribution mismatch in semi-supervised deep learning (SSDL) with MixMatch. This work is divided into two components: (i) an extensive…

Multi-class ClassificationSemantic SimilaritySemantic Textual Similarity