paper-with-me

홈 › Papers

Reinforcement Learning for Reachability: Guaranteeing Asymptotic Optimality

2026-05-23 · Amogh Palasamudram, Jakub Svoboda, Suguman Bansal, Krishnendu Chatterjee arxiv

Reinforcement learning (RL) for reachability specifications is fundamental in sequential decision-making, yet theoretical guarantees remain less explored. A recent work achieves asymptotic convergence to optimal policies. However, this approach provides limited insight into convergence dynamics. In this work, we present an alternative approach that provides deeper theoretical insights into convergence. Our approach builds on PAC learning with assumptions. PAC learning guarantees near-optimal policies with high confidence in finite time but requires knowing internal MDP parameters like minimum transition probability. We argue that while these parameters are unknown in RL, they can be iteratively refined and estimated with increasing accuracy. By iteratively satisfying PAC conditions, we show that exact optimality can be achieved in the limit. Empirical evaluations on standard benchmarks validate our theoretical insights into convergence dynamics.

📄 PDF Abstract BibTeX arXiv:2605.24740

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning

Similar Papers 제목 키워드 기반

Policy Optimization with Linear Temporal Logic Constraints

2022-06-20 · Cameron Voloshin, Hoang M. Le, Swarat Chaudhuri, Yisong Yue

We study the problem of policy optimization (PO) with linear temporal logic (LTL) constraints. The language of LTL allows flexible description of tasks that may be unnatural to encode as a scalar cost function. We consid…

A Forward Reachability Perspective on Robust Control Invariance and Discount Factors in Reachability Analysis

2023-10-26 · Jason J. Choi, Donggun Lee, Boyang Li, Jonathan P. How 외

Control invariant sets are crucial for various methods that aim to design safe control policies for systems whose state constraints must be satisfied over an indefinite time horizon. In this article, we explore the conne…

valid

Iterative Reachability Estimation for Safe Reinforcement Learning

2023-09-24 · NeurIPS 2023 11

Ensuring safety is important for the practical deployment of reinforcement learning (RL). Various challenges must be addressed, such as handling stochasticity in the environments, providing rigorous guarantees of persist…

MuJoCoreinforcement-learningReinforcement LearningReinforcement Learning (RL)+1

Data-Driven Reachability Analysis Using Matrix Zonotopes

2020-11-17 · Amr Alanwar, Anne Koch, Frank Allgöwer, Karl Henrik Johansson

In this paper, we propose a data-driven reachability analysis approach for unknown system dynamics. Reachability analysis is an essential tool for guaranteeing safety properties. However, most current reachability analys…

An Optimal Procedure to Check Pareto-Optimality in House Markets with Single-Peaked Preferences

2020-02-14 · Aurélie Beynier, Nicolas Maudet, Simon Rey, Parham Shams

Recently, the problem of allocating one resource per agent with initial endowments (house markets) has seen a renewed interest: indeed, while in the domain of strict preferences the Top Trading Cycle algorithm is known t…