paper-with-me

홈 › Papers

Adaptive Experimental Design for Policy Learning

2024-01-08 · Masahiro Kato, Kyohei Okumura, Takuya Ishihara, Toru Kitagawa

This study investigates the contextual best arm identification (BAI) problem, aiming to design an adaptive experiment to identify the best treatment arm conditioned on contextual information (covariates). We consider a decision-maker who assigns treatment arms to experimental units during an experiment and recommends the estimated best treatment arm based on the contexts at the end of the experiment. The decision-maker uses a policy for recommendations, which is a function that provides the estimated best treatment arm given the contexts. In our evaluation, we focus on the worst-case expected regret, a relative measure between the expected outcomes of an optimal policy and our proposed policy. We derive a lower bound for the expected simple regret and then propose a strategy called Adaptive Sampling-Policy Learning (PLAS). We prove that this strategy is minimax rate-optimal in the sense that its leading factor in the regret upper bound matches the lower bound as the number of experimental units increases.

📄 PDF Abstract BibTeX arXiv:2401.03756

Code (0)

등록된 구현이 없습니다.

Tasks

counterfactualExperimental Design

Similar Papers 제목 키워드 기반

Implicit Deep Adaptive Design: Policy-Based Experimental Design without Likelihoods

2021-11-03 · NeurIPS 2021 12 · Desi R. Ivanova, Adam Foster, Steven Kleinegesse, Michael U. Gutmann 외

We introduce implicit Deep Adaptive Design (iDAD), a new method for performing adaptive experiments in real-time with implicit models. iDAD amortizes the cost of Bayesian optimal experimental design (BOED) by learning a …

Experimental Design

Step-DAD: Semi-Amortized Policy-Based Bayesian Experimental Design

2025-07-18 · Marcel Hedman, Desi R. Ivanova, Cong Guan, Tom Rainforth arxiv

We develop a semi-amortized, policy-based, approach to Bayesian experimental design (BED) called Stepwise Deep Adaptive Design (Step-DAD). Like existing, fully amortized, policy-based BED approaches, Step-DAD trains a de…

Test-time Adaptation

Adaptivity in Adaptive Submodularity

2019-11-09 · Hossein Esfandiari, Amin Karbasi, Vahab Mirrokni

Adaptive sequential decision making is one of the central challenges in machine learning and artificial intelligence. In such problems, the goal is to design an interactive policy that plans for an action to take, from a…

Active LearningDecision MakingExperimental DesignSequential Decision Making

JADAI: Jointly Amortizing Adaptive Design and Bayesian Inference

2025-12-28 · Niels Bracher, Lars Kühmichel, Desi R. Ivanova, Xavier Intes 외 arxiv

We consider problems of parameter estimation where design variables can be actively optimized to maximize information gain. To this end, we introduce JADAI, a framework that jointly amortizes Bayesian adaptive design and…

Bayesian Inference

Deep Adaptive Design: Amortizing Sequential Bayesian Experimental Design

2021-03-03 · Adam Foster, Desi R. Ivanova, Ilyas Malik, Tom Rainforth

We introduce Deep Adaptive Design (DAD), a method for amortizing the cost of adaptive Bayesian experimental design that allows experiments to be run in real-time. Traditional sequential Bayesian optimal experimental desi…

Experimental Design