paper-with-me

Papers

A Contraction Approach to Model-based Reinforcement Learning

2020-09-18 · Ting-Han Fan, Peter J. Ramadge

Despite its experimental success, Model-based Reinforcement Learning still lacks a complete theoretical understanding. To this end, we analyze the error in the cumulative reward using a contraction approach. We consider both stochastic and deterministic state transitions for continuous (non-discrete) state and action spaces. This approach doesn't require strong assumptions and can recover the typical quadratic error to the horizon. We prove that branched rollouts can reduce this error and are essential for deterministic transitions to have a Bellman contraction. Our analysis of policy mismatch error also applies to Imitation Learning. In this case, we show that GAN-type learning has an advantage over Behavioral Cloning when its discriminator is well-trained.

📄 PDF Abstract BibTeX arXiv:2009.08586

Code (0)

등록된 구현이 없습니다.

Tasks

Imitation LearningmodelModel-based Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

ContractionPPO: Certified Reinforcement Learning via Differentiable Contraction Layers

2026-03-20 · Vrushabh Zinage, Narek Harutyunyan, Eric Verheyden, Fred Y. Hadaegh 외 arxiv

Legged locomotion in unstructured environments demands not only high-performance control policies but also formal guarantees to ensure robustness under perturbations. Control methods often require carefully designed refe…

Reinforcement Learning

LOCO: Adaptive exploration in reinforcement learning via local estimation of contraction coefficients

2021-03-09 · ICLR Workshop SSL-RL 2021 5 · Manfred Diaz, Liam Paull, Pablo Samuel Castro

We offer a novel approach to balance exploration and exploitation in reinforcement learning (RL). To do so, we characterize an environment’s exploration difficulty via the Second Largest Eigenvalue Modulus (SLEM) of the …

reinforcement-learningReinforcement Learning (RL)

Optimizing Tensor Network Contraction Using Reinforcement Learning

2022-04-18 · Eli A. Meirom, Haggai Maron, Shie Mannor, Gal Chechik

Quantum Computing (QC) stands to revolutionize computing, but is currently still limited. To develop and test quantum algorithms today, quantum circuits are often simulated on classical computers. Simulating a complex qu…

Combinatorial Optimizationreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Multivariate Distributional Reinforcement Learning Using Sliced Divergences

2026-05-29 · Baptiste Debes, Tinne Tuytelaars arxiv

Distributional reinforcement learning (DRL) models the full return distribution rather than expectations, but extending it to multivariate settings remains challenging. Many common metrics do not naturally generalize bey…

Reinforcement LearningAtari Games

How do trout regulate patterns of muscle contraction to optimize propulsive efficiency during steady swimming

2025-12-01 · Tao Li, Chunze Zhang, Weiwei Yao, Junzhao He 외 arxiv

Understanding efficient fish locomotion offers insights for biomechanics, fluid dynamics, and engineering. Traditional studies often miss the link between neuromuscular control and whole-body movement. To explore energy …

Reinforcement Learning