paper-with-me

Papers

Making Non-Stochastic Control (Almost) as Easy as Stochastic

2020-06-10 · NeurIPS 2020 12 · Max Simchowitz

Recent literature has made much progress in understanding \emph{online LQR}: a modern learning-theoretic take on the classical control problem in which a learner attempts to optimally control an unknown linear dynamical system with fully observed state, perturbed by i.i.d. Gaussian noise. It is now understood that the optimal regret on time horizon $T$ against the optimal control law scales as $\widetilde{\Theta}(\sqrt{T})$. In this paper, we show that the same regret rate (against a suitable benchmark) is attainable even in the considerably more general non-stochastic control model, where the system is driven by \emph{arbitrary adversarial} noise (Agarwal et al. 2019). In other words, \emph{stochasticity confers little benefit in online LQR}. We attain the optimal $\widetilde{\mathcal{O}}(\sqrt{T})$ regret when the dynamics are unknown to the learner, and $\mathrm{poly}(\log T)$ regret when known, provided that the cost functions are strongly convex (as in LQR). Our algorithm is based on a novel variant of online Newton step (Hazan et al. 2007), which adapts to the geometry induced by possibly adversarial disturbances, and our analysis hinges on generic "policy regret" bounds for certain structured losses in the OCO-with-memory framework (Anava et al. 2015). Moreover, our results accomodate the full generality of the non-stochastic control setting: adversarially chosen (possibly non-quadratic) costs, partial state observation, and fully adversarial process and observation noise.

📄 PDF Abstract BibTeX arXiv:2006.05910

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

High Confidence Level Inference is Almost Free using Parallel Stochastic Optimization

2024-01-17 · Wanrong Zhu, Zhipeng Lou, Ziyang Wei, Wei Biao Wu

Uncertainty quantification for estimation through stochastic optimization solutions in an online setting has gained popularity recently. This paper introduces a novel inference method focused on constructing confidence i…

Stochastic OptimizationUncertainty Quantification

Solving Stochastic Compositional Optimization is Nearly as Easy as Solving Stochastic Optimization

2020-08-25 · Tianyi Chen, Yuejiao Sun, Wotao Yin

Stochastic compositional optimization generalizes classic (non-compositional) stochastic optimization to the minimization of compositions of functions. Each composition may introduce an additional expectation. The series…

ManagementMeta-LearningStochastic Optimization

Almost Dominance: Inference and Application

2023-12-04 · Xiaojun Song, Zhenting Sun

This paper proposes a general framework for inference on three types of almost dominances: Almost Lorenz dominance, almost inverse stochastic dominance, and almost stochastic dominance. We first generalize almost Lorenz …

Safe Stabilization for Stochastic Time-Delay Systems

2022-11-21 · Zhuo-Rui Pan, Wei Ren, Xi-Ming Sun

This paper addresses the safe stabilization problem of stochastic nonlinear time-delay systems. Based on theKrasovskii approach, we first propose a stochastic control Lyapunov-Krasovskii functional to guarantee the stabi…

Stability Verification in Stochastic Control Systems via Neural Network Supermartingales

2021-12-17 · Mathias Lechner, Đorđe Žikelić, Krishnendu Chatterjee, Thomas A. Henzinger

We consider the problem of formally verifying almost-sure (a.s.) asymptotic stability in discrete-time nonlinear stochastic control systems. While verifying stability in deterministic control systems is extensively studi…