paper-with-me

Papers

NeoRL: Efficient Exploration for Nonepisodic RL

2024-06-03 · Bhavya Sukhija, Lenart Treven, Florian Dörfler, Stelian Coros, Andreas Krause

We study the problem of nonepisodic reinforcement learning (RL) for nonlinear dynamical systems, where the system dynamics are unknown and the RL agent has to learn from a single trajectory, i.e., without resets. We propose Nonepisodic Optimistic RL (NeoRL), an approach based on the principle of optimism in the face of uncertainty. NeoRL uses well-calibrated probabilistic models and plans optimistically w.r.t. the epistemic uncertainty about the unknown dynamics. Under continuity and bounded energy assumptions on the system, we provide a first-of-its-kind regret bound of $\setO(\beta_T \sqrt{T \Gamma_T})$ for general nonlinear systems with Gaussian process dynamics. We compare NeoRL to other baselines on several deep RL environments and empirically demonstrate that NeoRL achieves the optimal average cost while incurring the least regret.

📄 PDF Abstract BibTeX arXiv:2406.01175

Code (0)

등록된 구현이 없습니다.

Tasks

Efficient ExplorationReinforcement Learning (RL)

Methods 이 논문이 사용한 방법론

Gaussian Process Gaussian Processes are non-parametric models for approximating functions. They rely upon a measure of similarity between points (the kernel function) to predict the value for…

Similar Papers 제목 키워드 기반

Towards neoRL networks; the emergence of purposive graphs

2022-02-25 · Per R. Leikanger

The neoRL framework for purposive AI implements latent learning by emulated cognitive maps, with general value functions (GVF) expressing operant desires toward separate states. The agent's expectancy of reward, expresse…

Navigate

NEORL: NeuroEvolution Optimization with Reinforcement Learning

2021-12-01 · Majdi I. Radaideh, Katelin Du, Paul Seurin, Devin Seyler 외

We present an open-source Python framework for NeuroEvolution Optimization with Reinforcement Learning (NEORL) developed at the Massachusetts Institute of Technology. NEORL offers a global optimization interface of state…

Benchmarkingglobal-optimizationreinforcement-learningReinforcement Learning+1

NeoRL: A Near Real-World Benchmark for Offline Reinforcement Learning

2021-02-01 · Rongjun Qin, Songyi Gao, Xingyuan Zhang, Zhen Xu 외

Offline reinforcement learning (RL) aims at learning a good policy from a batch of collected data, without extra interactions with the environment during training. However, current offline RL benchmarks commonly have a l…

Offline RLreinforcement-learningReinforcement LearningReinforcement Learning (RL)

NeoRL-2: Near Real-World Benchmarks for Offline Reinforcement Learning with Extended Realistic Scenarios

2025-03-25 · Songyi Gao, Zuolin Tu, Rong-Jun Qin, Yi-Hao Sun 외

Offline reinforcement learning (RL) aims to learn from historical data without requiring (costly) access to the environment. To facilitate offline RL research, we previously introduced NeoRL, which highlighted that datas…

BenchmarkingOffline RLreinforcement-learningReinforcement Learning+1

Navigating Conceptual Space; A new take on Artificial General Intelligence

2022-02-19 · Per R. Leikanger

Edward C. Tolman found reinforcement learning unsatisfactory for explaining intelligence and proposed a clear distinction between learning and behavior. Tolman's ideas on latent learning and cognitive maps eventually led…

Autonomous NavigationRobot Navigationvalid