paper-with-me

Papers

Model-based Offline Quantum Reinforcement Learning

2024-04-14 · Simon Eisenmann, Daniel Hein, Steffen Udluft, Thomas A. Runkler

This paper presents the first algorithm for model-based offline quantum reinforcement learning and demonstrates its functionality on the cart-pole benchmark. The model and the policy to be optimized are each implemented as variational quantum circuits. The model is trained by gradient descent to fit a pre-recorded data set. The policy is optimized with a gradient-free optimization scheme using the return estimate given by the model as the fitness function. This model-based approach allows, in principle, full realization on a quantum computer during the optimization phase and gives hope that a quantum advantage can be achieved as soon as sufficiently powerful quantum computers are available.

📄 PDF Abstract BibTeX arXiv:2404.10017

Code (0)

등록된 구현이 없습니다.

Tasks

modelreinforcement-learningReinforcement Learning

Similar Papers 제목 키워드 기반

Improved Offline Reinforcement Learning via Quantum Metric Encoding

2025-11-13 · Outongyi Lv, Yewei Yuan, Nana Liu arxiv

Reinforcement learning (RL) with limited samples is common in real-world applications. However, offline RL performance under this constraint is often suboptimal. We consider an alternative approach to dealing with limite…

Reinforcement LearningOffline RL

Quantum Decision Transformers (QDT): Synergistic Entanglement and Interference for Offline Reinforcement Learning

2025-12-09 · Abraham Itzhak Weinberg arxiv

Offline reinforcement learning enables policy learning from pre-collected datasets without environment interaction, but existing Decision Transformer (DT) architectures struggle with long-horizon credit assignment and co…

Reinforcement LearningContinuous Control

A Bit of Freedom Goes a Long Way: Classical and Quantum Algorithms for Reinforcement Learning under a Generative Model

2025-07-30 · Andris Ambainis, Joao F. Doriguello, Debbie Lim arxiv

We propose novel classical and quantum online algorithms for learning finite- and infinite-horizon Markov Decision Processes (MDPs). Our algorithms are based on a hybrid online-offline reinforcement learning model wherei…

Reinforcement Learning

Challenges in Applying Variational Quantum Algorithms to Dynamic Satellite Network Routing

2025-08-06 · Phuc Hao Do, Tran Duc Le arxiv

Applying near-term variational quantum algorithms to the problem of dynamic satellite network routing represents a promising direction for quantum computing. In this work, we provide a critical evaluation of two major ap…

Reinforcement Learning

Variational Quantum Circuits in Offline Contextual Bandit Problems

2025-09-09 · Lukas Schulte, Daniel Hein, Steffen Udluft, Thomas A. Runkler arxiv

This paper explores the application of variational quantum circuits (VQCs) for solving offline contextual bandit problems in industrial optimization tasks. Using the Industrial Benchmark (IB) environment, we evaluate the…