paper-with-me

Papers

Improved Model-based Reinforcement Learning with Smooth Kernels

2026-05-08 · Kun Long, Yuqiang Li, Xianyi Wu arxiv

For continuous state-action space scenarios, classical reinforcement learning (RL) theory predominantly focuses on low-rank Markov decision processes (MDPs), which provide sample-efficient guarantees at the expense of restrictive structural assumptions. Kernel smoothing model-based approaches offer a promising alternative paradigm that instead leverages the smoothness of the MDP and employs non-parametric kernel smoothing estimates of transition dynamics. This paper proposes a new kernel-smoothing model-based approach for online reinforcement learning in finite-horizon settings under Lipschitz continuity assumptions on the MDP. By incorporating a Bernstein-style exploration bonus into the kernel smoothing framework, our method achieves a regret bound which improves upon the state-of-the-art regret bound in its dependence on the horizon. The theoretical advancement relies on a delicate analysis of the synergy between Bernstein-style bonuses and kernel smoothing, where a new tight Bernstein-type concentration inequality for martingales may be of independent interest.

📄 PDF Abstract BibTeX arXiv:2605.07218

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning

Similar Papers 제목 키워드 기반

Probabilistic Smoothing with Ratio-Monotone Transforms for Global Optimization

2026-05-26 · Kukyoung Jang, Taehyun Cho, Junrui Zhang, Ping Xu 외 arxiv

Probabilistic smoothing is a standard tool for global optimization, but existing methods rely on Gaussian kernels and specific transforms, often resulting in strong hyperparameter sensitivity and limited robustness. We p…

Smooth Kernels Improve Adversarial Robustness and Perceptually-Aligned Gradients

2020-01-01 · ICLR 2020 1 · Haohan Wang, Xindi Wu, Songwei Ge, Zachary C. Lipton 외

Recent research has shown that CNNs are often overly sensitive to high-frequency textural patterns. Inspired by the intuition that humans are more sensitive to the lower-frequency (larger-scale) patterns we design a regu…

Adversarial Robustness

A Structural Smoothing Framework For Robust Graph Comparison

2015-12-01 · NeurIPS 2015 12 · Pinar Yanardag, S. V. N. Vishwanathan

In this paper, we propose a general smoothing framework for graph kernels by taking \textit{structural similarity} into account, and apply it to derive smoothed variants of popular graph kernels. Our framework is inspi…

Generalized Kernel Thinning

2021-10-04 · ICLR 2022 4 · Raaz Dwivedi, Lester Mackey

The kernel thinning (KT) algorithm of Dwivedi and Mackey (2021) compresses a probability distribution more effectively than independent sampling by targeting a reproducing kernel Hilbert space (RKHS) and leveraging a les…

Kernel-Based Reinforcement Learning: A Finite-Time Analysis

2020-04-12 · Omar Darwiche Domingues, Pierre Ménard, Matteo Pirotta, Emilie Kaufmann 외

We consider the exploration-exploitation dilemma in finite-horizon reinforcement learning problems whose state-action space is endowed with a metric. We introduce Kernel-UCBVI, a model-based optimistic algorithm that lev…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)