paper-with-me

Papers

Quantum Bandits

2020-02-15 · Balthazar Casalé, Giuseppe Di Molfetta, Hachem Kadri, Liva Ralaivola

We consider the quantum version of the bandit problem known as {\em best arm identification} (BAI). We first propose a quantum modeling of the BAI problem, which assumes that both the learning agent and the environment are quantum; we then propose an algorithm based on quantum amplitude amplification to solve BAI. We formally analyze the behavior of the algorithm on all instances of the problem and we show, in particular, that it is able to get the optimal solution quadratically faster than what is known to hold in the classical case.

📄 PDF Abstract BibTeX arXiv:2002.06395

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Quantum Heavy-tailed Bandits

2023-01-23 · Yulian Wu, Chaowen Guan, Vaneet Aggarwal, Di Wang

In this paper, we study multi-armed bandits (MAB) and stochastic linear bandits (SLB) with heavy-tailed rewards and quantum reward oracle. Unlike the previous work on quantum bandits that assumes bounded/sub-Gaussian dis…

Multi-Armed Bandits

Quantum Multi-Armed Bandits and Stochastic Linear Bandits Enjoy Logarithmic Regrets

2022-05-30 · Zongqi Wan, Zhijie Zhang, Tongyang Li, Jialin Zhang 외

Multi-arm bandit (MAB) and stochastic linear bandit (SLB) are important models in reinforcement learning, and it is well-known that classical algorithms for bandits with time horizon $T$ suffer $\Omega(\sqrt{T})$ regret.…

Multi-Armed Banditsreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Towards Noise-Resilient Quantum Multi-Armed and Stochastic Linear Bandits

2026-03-19 · Zhuoyue Chen, Kechao Cai arxiv

Quantum multi-armed bandits (MAB) and stochastic linear bandits (SLB) have recently attracted significant attention, as their quantum counterparts can achieve quadratic speedups over classical MAB and SLB. However, most …

Multi-Armed Bandits

Quantum Speedups of Optimizing Approximately Convex Functions with Applications to Logarithmic Regret Stochastic Convex Bandits

2022-09-26 · Tongyang Li, Ruizhe Zhang

We initiate the study of quantum algorithms for optimizing approximately convex functions. Given a convex set ${\cal K}\subseteq\mathbb{R}^{n}$ and a function $F\colon\mathbb{R}^{n}\to\mathbb{R}$ such that there exists a…

Quantum Algorithms for Bandits with Knapsacks with Improved Regret and Time Complexities

2025-07-06 · Yuexin Su, Ziyi Yang, Peiyuan Huang, Tongyang Li 외 arxiv

Bandits with knapsacks (BwK) constitute a fundamental model that combines aspects of stochastic integer programming with online learning. Classical algorithms for BwK with a time horizon $T$ achieve a problem-independent…

Multi-Armed Bandits