paper-with-me

홈 › Papers

Scalable Quantum Reinforcement Learning on NISQ Devices with Dynamic-Circuit Qubit Reuse and Grover Optimization

2025-09-19 · Thet Htar Su, Shaswot Shresthamali, Masaaki Kondo arxiv

A scalable and resource-efficient quantum reinforcement learning framework is presented that eliminates the linear qubit-scaling barrier in multi-step quantum Markov decision processes (QMDPs). The proposed framework integrates a QMDP formulation, dynamic-circuit execution, and Grover-based amplitude amplification into a unified quantum-native architecture. Environment dynamics are encoded entirely within quantum Hilbert space, enabling coherent superposition over state-action sequences and a direct quantum agent-environment interface without intermediate quantum-to-classical conversion. The central contribution is a dynamic execution model for multi-step QMDPs that employs mid-circuit measurement and reset to recycle a fixed physical quantum register across sequential interactions. This approach preserves trajectory fidelity relative to a static unrolled QMDP, generating identical state-action sequences while reducing the physical qubit requirement from 7xT to a constant 7, independent of the interaction horizon T. Thus, the qubit complexity of multi-step QMDPs is transformed from O(T) to O(1) while maintaining functional equivalence at the level of trajectory generation. Trajectory returns are evaluated via quantum arithmetic, and high-return trajectories are marked and amplified using amplitude amplification to increase their sampling probability. Simulations confirm preservation of trajectory fidelity with a 66% qubit reduction compared to a static design. Experimental execution on an IBM Heron-class processor demonstrates feasibility on noisy intermediate-scale quantum hardware, establishing a scalable and resource-efficient foundation for large-scale quantum-native reinforcement learning.

📄 PDF Abstract BibTeX arXiv:2509.16002

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning

Similar Papers 제목 키워드 기반

Machine-learning based noise characterization and correction on neutral atoms NISQ devices

2023-06-27 · Ettore Canonici, Stefano Martina, Riccardo Mengoni, Daniele Ottaviani 외

Neutral atoms devices represent a promising technology that uses optical tweezers to geometrically arrange atoms and modulated laser pulses to control the quantum states. A neutral atoms Noisy Intermediate Scale Quantum …

Reinforcement Learning (RL)

Gated QKAN-FWP: Scalable Quantum-inspired Sequence Learning

2026-05-07 · Kuo-Chung Peng, Samuel Yen-Chi Chen, Jiun-Cheng Jiang, Chen-Yu Liu 외 arxiv

Fast Weight Programmers (FWPs) encode temporal dependencies through dynamically updated parameters rather than recurrent hidden states. Quantum FWPs (QFWPs) extend this idea with variational quantum circuits (VQCs), but …

Reinforcement Learning

Quantum-Assisted Clustering Algorithms for NISQ-Era Devices

2019-04-18 · Samuel S. Mendelson, Robert W. Strand, Guy B. Oldaker IV, Jacob M. Farinholt

In the NISQ-era of quantum computing, we should not expect to see quantum devices that provide an exponential improvement in runtime for practical problems, due to the lack of error correction and small number of qubits …

Clustering

Dynamic Estimation Loss Control in Variational Quantum Sensing via Online Conformal Inference

2025-05-29 · Ivana Nikoloska, Hamdi Joudeh, Ruud Van Sloun, Osvaldo Simeone

Quantum sensing exploits non-classical effects to overcome limitations of classical sensors, with applications ranging from gravitational-wave detection to nanoscale imaging. However, practical quantum sensors built on n…

Gravitational Wave Detection

Variational Quantum Circuits for Deep Reinforcement Learning

2019-06-30 · Samuel Yen-Chi Chen, Chao-Han Huck Yang, Jun Qi, Pin-Yu Chen 외

The state-of-the-art machine learning approaches are based on classical von Neumann computing architectures and have been widely used in many industrial and academic domains. With the recent development of quantum comput…

BIG-bench Machine LearningDecision MakingDeep Reinforcement LearningQuantum Machine Learning+3