paper-with-me

홈 › Papers

Sharp analysis of linear ensemble sampling

2026-02-08 · David Janz, Arya Akhavan, Csaba Szepesvári arxiv

We analyse linear ensemble sampling (ES) with standard Gaussian perturbations in stochastic linear bandits. We show that for ensemble size $m=Θ(d\log n)$, ES attains $\tilde O(d^{3/2}\sqrt n)$ high-probability regret, closing the gap to the Thompson sampling benchmark while keeping computation comparable. The proof brings a new perspective on randomized exploration in linear bandits by reducing the analysis to a time-uniform exceedance problem for $m$ independent Brownian motions. This continuous-time lens appears particularly natural here: it yields an exact representation of the relevant discrete-time processes, and we do not know another route to a sharp ES bound.

📄 PDF Abstract BibTeX arXiv:2602.08026

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Ensemble sampling for linear bandits: small ensembles suffice

2023-11-14 · David Janz, Alexander E. Litvak, Csaba Szepesvári

We provide the first useful and rigorous analysis of ensemble sampling for the stochastic linear bandit setting. In particular, we show that, under standard assumptions, for a $d$-dimensional stochastic linear bandit wit…

Improved Regret of Linear Ensemble Sampling

2024-11-06 · Harin Lee, Min-hwan Oh

In this work, we close the fundamental gap of theory and practice by providing an improved regret bound for linear ensemble sampling. We prove that with an ensemble size logarithmic in $T$, linear ensemble sampling can a…

An Analysis of Ensemble Sampling

2022-03-02 · Chao Qin, Zheng Wen, Xiuyuan Lu, Benjamin Van Roy

Ensemble sampling serves as a practical approximation to Thompson sampling when maintaining an exact posterior distribution over model parameters is computationally intractable. In this paper, we establish a regret bound…

Thompson Sampling

Sharpness-diversity tradeoff: improving flat ensembles with SharpBalance

2024-07-17 · Haiquan Lu, Xiaotian Liu, Yefan Zhou, Qunli Li 외

Recent studies on deep ensembles have identified the sharpness of the local minima of individual learners and the diversity of the ensemble members as key factors in improving test-time performance. Building on this, our…

Diversity

Provable Anytime Ensemble Sampling Algorithms in Nonlinear Contextual Bandits

2025-10-12 · Jiazheng Sun, Weixin Wang, Pan Xu arxiv

We provide a unified algorithmic framework for ensemble sampling in nonlinear contextual bandits and develop corresponding regret bounds for two most common nonlinear contextual bandit settings: Generalized Linear Ensemb…