paper-with-me

Papers

On Bayesian index policies for sequential resource allocation

2016-01-06 · Emilie Kaufmann

This paper is about index policies for minimizing (frequentist) regret in a stochastic multi-armed bandit model, inspired by a Bayesian view on the problem. Our main contribution is to prove that the Bayes-UCB algorithm, which relies on quantiles of posterior distributions, is asymptotically optimal when the reward distributions belong to a one-dimensional exponential family, for a large class of prior distributions. We also show that the Bayesian literature gives new insight on what kind of exploration rates could be used in frequentist, UCB-type algorithms. Indeed, approximations of the Bayesian optimal solution or the Finite Horizon Gittins indices provide a justification for the kl-UCB+ and kl-UCB-H+ algorithms, whose asymptotic optimality is also established.

📄 PDF Abstract BibTeX arXiv:1601.01190

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Possible and Necessary Allocations via Sequential Mechanisms

2014-12-06 · Haris Aziz, Toby Walsh, Lirong Xia

A simple mechanism for allocating indivisible resources is sequential allocation in which agents take turns to pick items. We focus on possible and necessary allocation problems, checking whether allocations of a given f…

Fair Resource Allocation in Weakly Coupled Markov Decision Processes

2024-11-14 · Xiaohui Tu, Yossiri Adulyasak, Nima Akbarzadeh, Erick Delage

We consider fair resource allocation in sequential decision-making environments modeled as weakly coupled Markov decision processes, where resource constraints couple the action spaces of $N$ sub-Markov decision processe…

Decision MakingDeep Reinforcement LearningFairnessSequential Decision Making

Evaluating the Effectiveness of Index-Based Treatment Allocation

2024-02-19 · Niclas Boehmer, Yash Nair, Sanket Shah, Lucas Janson 외

When resources are scarce, an allocation policy is needed to decide who receives a resource. This problem occurs, for instance, when allocating scarce medical resources and is often solved using modern ML methods. This p…

valid

Transformers as Bayesian In-Context Experimenters: Smoothness-Adaptive Efficient ATE Estimation

2026-06-30 · Jiachun Li, David Simchi-Levi arxiv

Adaptive experiments for average treatment effects (ATE) require randomized allocations balancing valid inference with statistical efficiency. The oracle design is a covariate-dependent Neyman rule governed by unknown ar…

Bayes-Optimal Effort Allocation in Crowdsourcing: Bounds and Index Policies

2015-12-31 · Weici Hu, Peter I. Frazier

We consider effort allocation in crowdsourcing, where we wish to assign labeling tasks to imperfect homogeneous crowd workers to maximize overall accuracy in a continuous-time Bayesian setting, subject to budget and time…