paper-with-me

홈 › Papers

Simulation-Based Inference for Adaptive Experiments

2025-06-03 · Brian M Cho, Aurélien Bibaut, Nathan Kallus

Multi-arm bandit experimental designs are increasingly being adopted over standard randomized trials due to their potential to improve outcomes for study participants, enable faster identification of the best-performing options, and/or enhance the precision of estimating key parameters. Current approaches for inference after adaptive sampling either rely on asymptotic normality under restricted experiment designs or underpowered martingale concentration inequalities that lead to weak power in practice. To bypass these limitations, we propose a simulation-based approach for conducting hypothesis tests and constructing confidence intervals for arm specific means and their differences. Our simulation-based approach uses positively biased nuisances to generate additional trajectories of the experiment, which we call \textit{simulation with optimism}. Using these simulations, we characterize the distribution potentially non-normal sample mean test statistic to conduct inference. We provide guarantees for (i) asymptotic type I error control, (ii) convergence of our confidence intervals, and (iii) asymptotic strong consistency of our estimator over a wide variety of common bandit designs. Our empirical results show that our approach achieves the desired coverage while reducing confidence interval widths by up to 50%, with drastic improvements for arms not targeted by the design.

📄 PDF Abstract BibTeX arXiv:2506.02881

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Inference for Batched Adaptive Experiments

2025-12-10 · Jan Kemper, Davud Rostam-Afschar arxiv

The advantages of adaptive experiments have led to their rapid adoption in economics, other fields, as well as among practitioners. However, adaptive experiments pose challenges for causal inference. This note suggests a…

Causal Inference

BayesSimIG: Scalable Parameter Inference for Adaptive Domain Randomization with IsaacGym

2021-07-09 · Rika Antonova, Fabio Ramos, Rafael Possas, Dieter Fox

BayesSim is a statistical technique for domain randomization in reinforcement learning based on likelihood-free inference of simulation parameters. This paper outlines BayesSimIG: a library that provides an implementatio…

GPUReinforcement Learning (RL)

Dynamic SBI: Round-free Sequential Simulation-Based Inference with Adaptive Datasets

2025-10-15 · Huifang Lyu, James Alvey, Noemi Anau Montel, Mauro Pieroni 외 arxiv

Simulation-based inference (SBI) is emerging as a new statistical paradigm for addressing complex scientific inference problems. By leveraging the representational power of deep neural networks, SBI can extract the most …

Hybrid-Adaptive Thread Tuning to Mitigate Simulation Execution Bottlenecks in High-Performance Reinforcement Learning Inference

2026-08-06 · Jiming Su, Hantao Hua, Lujia Yin, Yiping Yao 외 arxiv

In simulation-in-the-loop decision-making systems, reinforcement learning (RL) inference is often constrained by simulator-side execution overhead, where workloads are highly dynamic and sensitive to runtime thread confi…

Reinforcement Learning

The Adaptive Doubly Robust Estimator for Policy Evaluation in Adaptive Experiments and a Paradox Concerning Logging Policy

2020-10-08 · Masahiro Kato, Shota Yasui, Kenichiro McAlinn

The doubly robust (DR) estimator, which consists of two nuisance parameters, the conditional mean outcome and the logging policy (the probability of choosing an action), is crucial in causal inference. This paper propose…

Causal InferenceTime Series Analysis