paper-with-me

Papers

Stochastic Approximation with Decision-Dependent Distributions: Asymptotic Normality and Optimality

2022-07-09 · Joshua Cutler, Mateo Díaz, Dmitriy Drusvyatskiy

We analyze a stochastic approximation algorithm for decision-dependent problems, wherein the data distribution used by the algorithm evolves along the iterate sequence. The primary examples of such problems appear in performative prediction and its multiplayer extensions. We show that under mild assumptions, the deviation between the average iterate of the algorithm and the solution is asymptotically normal, with a covariance that clearly decouples the effects of the gradient noise and the distributional shift. Moreover, building on the work of H\'ajek and Le Cam, we show that the asymptotic performance of the algorithm with averaging is locally minimax optimal.

📄 PDF Abstract BibTeX arXiv:2207.04173

Code (1)

mateodd25/asymptotic-normality-in-performative-prediction 공식 구현

Similar Papers 제목 키워드 기반

Ranking and Selection as Stochastic Control

2017-10-07 · Yijie Peng, Edwin K. P. Chong, Chun-Hung Chen, Michael C. Fu

Under a Bayesian framework, we formulate the fully sequential sampling and selection decision in statistical ranking and selection as a stochastic control problem, and derive the associated Bellman equation. Using value …

Optimal variance-reduced stochastic approximation in Banach spaces

2022-01-21 · Wenlong Mou, Koulik Khamaru, Martin J. Wainwright, Peter L. Bartlett 외

We study the problem of estimating the fixed point of a contractive operator defined on a separable Banach space. Focusing on a stochastic query model that provides noisy evaluations of the operator, we analyze a varianc…

Q-Learning

Gaussian Approximation for Two-Timescale Linear Stochastic Approximation

2025-08-11 · Bogdan Butyrin, Artemy Rubtsov, Alexey Naumov, Vladimir Ulyanov 외 arxiv

In this paper, we establish non-asymptotic bounds for accuracy of normal approximation for linear two-timescale stochastic approximation (TTSA) algorithms driven by martingale difference or Markov noise. Focusing on both…

Is Temporal Difference Learning Optimal? An Instance-Dependent Analysis

2020-03-16 · Koulik Khamaru, Ashwin Pananjady, Feng Ruan, Martin J. Wainwright 외

We address the problem of policy evaluation in discounted Markov decision processes, and provide instance-dependent guarantees on the $\ell_\infty$-error under a generative model. We establish both asymptotic and non-asy…

Non-asymptotic approximations of Gaussian neural networks via second-order Poincaré inequalities

2023-04-08 · Alberto Bordino, Stefano Favaro, Sandra Fortini

There is a growing interest on large-width asymptotic properties of Gaussian neural networks (NNs), namely NNs whose weights are initialized according to Gaussian distributions. A well-established result is that, as the …