paper-with-me

홈 › Papers

Quantum Natural Policy Gradients: Towards Sample-Efficient Reinforcement Learning

2023-04-26 · Nico Meyer, Daniel D. Scherer, Axel Plinge, Christopher Mutschler, Michael J. Hartmann

Reinforcement learning is a growing field in AI with a lot of potential. Intelligent behavior is learned automatically through trial and error in interaction with the environment. However, this learning process is often costly. Using variational quantum circuits as function approximators potentially can reduce this cost. In order to implement this, we propose the quantum natural policy gradient (QNPG) algorithm -- a second-order gradient-based routine that takes advantage of an efficient approximation of the quantum Fisher information matrix. We experimentally demonstrate that QNPG outperforms first-order based training on Contextual Bandits environments regarding convergence speed and stability and moreover reduces the sample complexity. Furthermore, we provide evidence for the practical feasibility of our approach by training on a 12-qubit hardware device.

📄 PDF Abstract BibTeX arXiv:2304.13571

Code (1)

nicomeyer96/quantum-natural-policy-gradients 공식 구현

Tasks

Multi-Armed Banditsreinforcement-learningReinforcement Learning

Methods 이 논문이 사용한 방법론

SPEED The monocular depth estimation (MDE) is the task of estimating depth from a single frame. This information is an essential knowledge in many computer vision tasks such as scene…

Similar Papers 제목 키워드 기반

Policy Gradients using Variational Quantum Circuits

2022-03-20 · André Sequeira, Luis Paulo Santos, Luís Soares Barbosa

Variational Quantum Circuits are being used as versatile Quantum Machine Learning models. Some empirical results exhibit an advantage in supervised and generative learning tasks. However, when applied to Reinforcement Le…

BenchmarkingQuantum Machine Learningreinforcement-learningReinforcement Learning+1

Accelerating Quantum Reinforcement Learning with a Quantum Natural Policy Gradient Based Approach

2025-01-27 · Yang Xu, Vaneet Aggarwal

We address the problem of quantum reinforcement learning (QRL) under model-free settings with quantum oracle access to the Markov Decision Process (MDP). This paper introduces a Quantum Natural Policy Gradient (QNPG) alg…

Rank-1 Approximation of Inverse Fisher for Natural Policy Gradients in Deep Reinforcement Learning

2026-01-26 · Yingxiao Huo, Satya Prakash Dash, Radu Stoican, Samuel Kaski 외 arxiv

Natural gradients have long been studied in deep reinforcement learning due to their fast convergence properties and covariant weight updates. However, computing natural gradients requires inversion of the Fisher Informa…

Reinforcement Learning

Trainability issues in quantum policy gradients

2024-06-13 · André Sequeira, Luis Paulo Santos, Luis Soares Barbosa

This research explores the trainability of Parameterized Quantum circuit-based policies in Reinforcement Learning, an area that has recently seen a surge in empirical exploration. While some studies suggest improved samp…

On Quantum Natural Policy Gradients

2024-01-16 · André Sequeira, Luis Paulo Santos, Luis Soares Barbosa

This research delves into the role of the quantum Fisher Information Matrix (FIM) in enhancing the performance of Parameterized Quantum Circuit (PQC)-based reinforcement learning agents. While previous studies have highl…

Multi-Armed Banditsreinforcement-learningReinforcement Learning