paper-with-me

홈 › Papers

Adaptive Policy Learning Under Unknown Network Interference

2026-05-11 · Aidan Gleich, Eric Laber, Alexander Volfovsky arxiv

Adaptive experimentation under unknown network interference requires solving two coupled problems: (i) learning the underlying dynamics of interference among units and (ii) using these dynamics to inform treatment allocation in order to maximize a cumulative outcome of interest (e.g. revenue). Existing adaptive experimentation methods either assume the interference network is fully known or bypass the network by operating on coarse cluster-level randomizations. We develop a Thompson sampling algorithm that jointly learns the interference network and adaptively optimizes individual-level treatment allocations via a Gibbs sampler. The algorithm returns both an optimized treatment policy and an estimate of the interference network; the latter supports downstream causal analyses such as estimation of direct, indirect, and total treatment effects. For additive spillover models, we show that total reward is linear in the treatment vector with coefficients given by an $n$-dimensional latent score. We prove a Bayesian regret bound of order $\sqrt{nT \cdot B \log(en/B)}$ for exact posterior sampling; empirically, our Gibbs-based approximate sampler achieves regret consistent with this rate and remains sublinear when the additive spillovers assumption is violated. For general Neighborhood Interference, where this reduction is unavailable, we analyze an explore-then-commit variant with $O(n^2 \log T)$ graph-discovery cost. An information-theoretic $Ω(n \log T)$ lower bound complements both results. Empirically, our method achieves more than an order-of-magnitude reduction in regret in head-to-head comparisons. On two real-world networks, the algorithm achieves sublinear regret and yields downstream effect estimates with small RMSE relative to the truth.

📄 PDF Abstract BibTeX arXiv:2605.11191

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Joint ML-Bayesian Approach to Adaptive Radar Detection in the presence of Gaussian Interference

2025-03-04 · Chaoran Yin, Tianqi Wang, Linjie Yan, Chengpeng Hao 외

This paper addresses the adaptive radar target detection problem in the presence of Gaussian interference with unknown statistical properties. To this end, the problem is first formulated as a binary hypothesis test, and…

Neighborhood Adaptive Estimators for Causal Inference under Network Interference

2022-12-07 · Alexandre Belloni, Fei Fang, Alexander Volfovsky

Estimating causal effects has become an integral part of most applied fields. In this work we consider the violation of the classical no-interference assumption with units connected by a network. For tractability, we con…

Causal InferenceFeature Engineering

Policy design in experiments with unknown interference

2020-11-16 · Davide Viviano, Jess Rudder

This paper studies experimental designs for estimation and inference on policies with spillover effects. Units are organized into a finite number of large clusters and interact in unknown ways within each cluster. First,…

Experimental DesignTwo-sample testing

Bridging Adaptivity and Safety: Learning Agile Collision-Free Locomotion Across Varied Physics

2025-01-08 · Yichao Zhong, Chong Zhang, Tairan He, Guanya Shi

Real-world legged locomotion systems often need to reconcile agility and safety for different scenarios. Moreover, the underlying dynamics are often unknown and time-variant (e.g., payload, friction). In this paper, we i…

Friction

ReinWiFi: Application-Layer QoS Optimization of WiFi Networks with Reinforcement Learning

2024-05-06 · Qianren Li, Bojie Lv, Yuncong Hong, Rui Wang

The enhanced distributed channel access (EDCA) mechanism is used in current wireless fidelity (WiFi) networks to support priority requirements of heterogeneous applications. However, the EDCA mechanism can not adapt to p…

reinforcement-learningReinforcement LearningScheduling