paper-with-me

Papers

Minimax and Bayes Optimal Adaptive Experimental Design for Treatment Choice

2025-12-09 · Masahiro Kato arxiv

We consider an adaptive experiment for treatment choice and design a minimax and Bayes optimal adaptive experiment with respect to regret. Given binary treatments, the experimenter's goal is to choose the treatment with the highest expected outcome through an adaptive experiment, in order to maximize welfare. We consider adaptive experiments that consist of two phases, the treatment allocation phase and the treatment choice phase. The experiment starts with the treatment allocation phase, where the experimenter allocates treatments to experimental subjects to gather observations. During this phase, the experimenter can adaptively update the allocation probabilities using the observations obtained in the experiment. After the allocation phase, the experimenter proceeds to the treatment choice phase, where one of the treatments is selected as the best. For this adaptive experimental procedure, we propose an adaptive experiment that splits the treatment allocation phase into two stages, where we first estimate the standard deviations and then allocate each treatment proportionally to its standard deviation. We show that this experiment, often referred to as Neyman allocation, is minimax and Bayes optimal in the sense that its regret upper bounds exactly match the lower bounds that we derive. To show this optimality, we derive minimax and Bayes lower bounds for the regret using change-of-measure arguments. Then, we evaluate the corresponding upper bounds using the central limit theorem and large deviation bounds.

📄 PDF Abstract BibTeX arXiv:2512.08513

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Minimax and Bayes Optimal Best-Arm Identification

2025-06-30 · Masahiro Kato arxiv

This study investigates minimax and Bayes optimal strategies for fixed-budget best-arm identification. We consider an adaptive procedure consisting of a sampling phase followed by a recommendation phase. Within this fram…

Optimal Experiment Design for Causal Discovery from Fixed Number of Experiments

2017-02-27 · AmirEmad Ghassami, Saber Salehkaleybar, Negar Kiyavash

We study the problem of causal structure learning over a set of random variables when the experimenter is allowed to perform at most $M$ experiments in a non-adaptive manner. We consider the optimal learning strategy in …

Causal Discovery

On Uninformative Optimal Policies in Adaptive LQR with Unknown B-Matrix

2020-11-18 · Ingvar Ziemann, Henrik Sandberg

This paper presents local asymptotic minimax regret lower bounds for adaptive Linear Quadratic Regulators (LQR). We consider affinely parametrized $B$-matrices and known $A$-matrices and aim to understand when logarithmi…

Adaptive Experimental Design for Policy Learning

2024-01-08 · Masahiro Kato, Kyohei Okumura, Takuya Ishihara, Toru Kitagawa

This study investigates the contextual best arm identification (BAI) problem, aiming to design an adaptive experiment to identify the best treatment arm conditioned on contextual information (covariates). We consider a d…

counterfactualExperimental Design

Deep Adaptive Design: Amortizing Sequential Bayesian Experimental Design

2021-03-03 · Adam Foster, Desi R. Ivanova, Ilyas Malik, Tom Rainforth

We introduce Deep Adaptive Design (DAD), a method for amortizing the cost of adaptive Bayesian experimental design that allows experiments to be run in real-time. Traditional sequential Bayesian optimal experimental desi…

Experimental Design