paper-with-me

홈 › Papers

How to sample and when to stop sampling: The generalized Wald problem and minimax policies

2022-10-28 · Karun Adusumilli

We study sequential experiments where sampling is costly and a decision-maker aims to determine the best treatment for full scale implementation by (1) adaptively allocating units between two possible treatments, and (2) stopping the experiment when the expected welfare (inclusive of sampling costs) from implementing the chosen treatment is maximized. Working under a continuous time limit, we characterize the optimal policies under the minimax regret criterion. We show that the same policies also remain optimal under both parametric and non-parametric outcome distributions in an asymptotic regime where sampling costs approach zero. The minimax optimal sampling rule is just the Neyman allocation: it is independent of sampling costs and does not adapt to observed outcomes. The decision-maker halts sampling when the product of the average treatment difference and the number of observations surpasses a specific threshold. The results derived also apply to the so-called best-arm identification problem, where the number of observations is exogenously specified.

📄 PDF Abstract BibTeX arXiv:2210.15841

Code (0)

등록된 구현이 없습니다.

Tasks

Experimental Design

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

Prolonged Learning and Hasty Stopping: the Wald Problem with Ambiguity

2022-08-30 · Sarah Auster, Yeon-Koo Che, Konrad Mierendorff

This paper studies sequential information acquisition by an ambiguity-averse decision maker (DM), who decides how long to collect information before taking an irreversible action. The agent optimizes against the worst-ca…

Unified Precision-Guaranteed Stopping Rules for Contextual Learning

2026-04-09 · Mingrui Ding, Qiuhong Zhao, Siyang Gao, Jing Dong arxiv

Contextual learning seeks to learn a decision policy that maps an individual's characteristics to an action through data collection. In operations management, such data may come from various sources, and a central questi…

Early Stopping for Nonparametric Testing

2018-05-25 · NeurIPS 2018 12 · Meimei Liu, Guang Cheng

Early stopping of iterative algorithms is an algorithmic regularization method to avoid over-fitting in estimation and classification. In this paper, we show that early stopping can also be applied to obtain the minimax …

General Classification

Sequential Consensus for Multi-Agent LLM Debates: A Wald-SPRT compute governor with calibration-based failure detection

2026-05-18 · Andrea Morandi arxiv

Multi-agent LLM debate improves factuality and reasoning, but most recipes pick a fixed round count, over-spending on easy items and under-spending on hard ones. We adapt Wald's Sequential Probability Ratio Test (SPRT) a…

Wald-Kernel: Learning to Aggregate Information for Sequential Inference

2015-08-31 · Diyan Teng, Emre Ertin

Sequential hypothesis testing is a desirable decision making strategy in any time sensitive scenario. Compared with fixed sample-size testing, sequential testing is capable of achieving identical probability of error req…

Decision MakingTwo-sample testing