paper-with-me

Papers

Optimal oracle inequalities for solving projected fixed-point equations

2020-12-09 · Wenlong Mou, Ashwin Pananjady, Martin J. Wainwright

Linear fixed point equations in Hilbert spaces arise in a variety of settings, including reinforcement learning, and computational methods for solving differential and integral equations. We study methods that use a collection of random observations to compute approximate solutions by searching over a known low-dimensional subspace of the Hilbert space. First, we prove an instance-dependent upper bound on the mean-squared error for a linear stochastic approximation scheme that exploits Polyak--Ruppert averaging. This bound consists of two terms: an approximation error term with an instance-dependent approximation factor, and a statistical error term that captures the instance-specific complexity of the noise when projected onto the low-dimensional subspace. Using information theoretic methods, we also establish lower bounds showing that both of these terms cannot be improved, again in an instance-dependent sense. A concrete consequence of our characterization is that the optimal approximation factor in this problem can be much larger than a universal constant. We show how our results precisely characterize the error of a class of temporal difference learning methods for the policy evaluation problem with linear function approximation, establishing their optimality.

📄 PDF Abstract BibTeX arXiv:2012.05299

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Optimal Extragradient-Based Algorithms for Stochastic Variational Inequalities with Separable Structure

2023-09-21 · NeurIPS 2023 11

We consider the problem of solving stochastic monotone variational inequalities with a separable structure using a stochastic first-order oracle. Building on standard extragradient for variational inequalities we propose…

Forward-backward-forward methods with variance reduction for stochastic variational inequalities

2019-02-09 · Radu Ioan Bot, Panayotis Mertikopoulos, Mathias Staudigl, Phan Tu Vuong

We develop a new stochastic algorithm with variance reduction for solving pseudo-monotone stochastic variational inequalities. Our method builds on Tseng's forward-backward-forward (FBF) algorithm, which is known in the …

Aggregation of Affine Estimators

2013-11-12 · Dong Dai, Philippe Rigollet, Lucy Xia, Tong Zhang

We consider the problem of aggregating a general collection of affine estimators for fixed design regression. Relevant examples include some commonly used statistical estimators such as least squares, ridge and robust le…

Model Selection

Span-Agnostic Optimal Sample Complexity and Oracle Inequalities for Average-Reward RL

2025-02-16 · Matthew Zurek, Yudong Chen

We study the sample complexity of finding an $\varepsilon$-optimal policy in average-reward Markov Decision Processes (MDPs) with a generative model. The minimax optimal span-based complexity of $\widetilde{O}(SAH/\varep…

Analyzing the discrepancy principle for kernelized spectral filter learning algorithms

2020-04-17 · Alain Celisse, Martin Wahl

We investigate the construction of early stopping rules in the nonparametric regression problem where iterative learning algorithms are used and the optimal iteration number is unknown. More precisely, we study the discr…