paper-with-me

Papers

Online Statistical Inference for Nonlinear Stochastic Approximation with Markovian Data

2023-02-15 · Xiang Li, Jiadong Liang, Zhihua Zhang

We study the statistical inference of nonlinear stochastic approximation algorithms utilizing a single trajectory of Markovian data. Our methodology has practical applications in various scenarios, such as Stochastic Gradient Descent (SGD) on autoregressive data and asynchronous Q-Learning. By utilizing the standard stochastic approximation (SA) framework to estimate the target parameter, we establish a functional central limit theorem for its partial-sum process, $\boldsymbol{\phi}_T$. To further support this theory, we provide a matching semiparametric efficient lower bound and a non-asymptotic upper bound on its weak convergence, measured in the L\'evy-Prokhorov metric. This functional central limit theorem forms the basis for our inference method. By selecting any continuous scale-invariant functional $f$, the asymptotic pivotal statistic $f(\boldsymbol{\phi}_T)$ becomes accessible, allowing us to construct an asymptotically valid confidence interval. We analyze the rejection probability of a family of functionals $f_m$, indexed by $m \in \mathbb{N}$, through theoretical and numerical means. The simulation results demonstrate the validity and efficiency of our method.

📄 PDF Abstract BibTeX arXiv:2302.07690

Code (0)

등록된 구현이 없습니다.

Tasks

Q-Learningvalid

Methods 이 논문이 사용한 방법론

Q-Learning Q-Learning is an off-policy temporal difference control algorithm: $$Q\left(S\_{t}, A\_{t}\right) \leftarrow Q\left(S\_{t}, A\_{t}\right) + \alpha\left[R_{t+1} +…

Similar Papers 제목 키워드 기반

Efficient Stochastic Optimal Control through Approximate Bayesian Input Inference

2021-05-17 · Joe Watson, Hany Abdulsamad, Rolf Findeisen, Jan Peters

Optimal control under uncertainty is a prevailing challenge for many reasons. One of the critical difficulties lies in producing tractable solutions for the underlying stochastic optimization problem. We show how advance…

Stochastic Optimization

Stochastic Nonlinear Control via Finite-dimensional Spectral Dynamic Embedding

2023-04-08 · Zhaolin Ren, Tongzheng Ren, Haitong Ma, Na Li 외

This paper proposes an approach, Spectral Dynamics Embedding Control (SDEC), to optimal control for nonlinear stochastic systems. This method reveals an infinite-dimensional feature representation induced by the system's…

Statistical Inference of Constrained Stochastic Optimization via Sketched Sequential Quadratic Programming

2022-05-27 · Sen Na, Michael W. Mahoney

We consider online statistical inference of constrained stochastic nonlinear optimization problems. We apply the Stochastic Sequential Quadratic Programming (StoSQP) method to solve these problems, which can be regarded …

Second-order methodsStochastic Optimization

Asymptotic Time-Uniform Inference for Parameters in Averaged Stochastic Approximation

2024-10-19 · Chuhan Xie, Kaicheng Jin, Jiadong Liang, Zhihua Zhang

We study time-uniform statistical inference for parameters in stochastic approximation (SA), which encompasses a bunch of applications in optimization and machine learning. To that end, we analyze the almost-sure converg…

valid

Stochastic Variational Bayesian Inference for a Nonlinear Forward Model

2020-07-03 · Michael A. Chappell, Martin S. Craig, Mark W. Woolrich

Variational Bayes (VB) has been used to facilitate the calculation of the posterior distribution in the context of Bayesian inference of the parameters of nonlinear models from data. Previously an analytical formulation …

Bayesian Inference