paper-with-me

홈 › Papers

Safe Control with Minimal Regret

2022-03-01 · Andrea Martin, Luca Furieri, Florian Dörfler, John Lygeros, Giancarlo Ferrari-Trecate

As we move towards safety-critical cyber-physical systems that operate in non-stationary and uncertain environments, it becomes crucial to close the gap between classical optimal control algorithms and adaptive learning-based methods. In this paper, we present an efficient optimization-based approach for computing a finite-horizon robustly safe control policy that minimizes dynamic regret, in the sense of the loss relative to the optimal sequence of control actions selected in hindsight by a clairvoyant controller. By leveraging the system level synthesis framework (SLS), our method extends recent results on regret minimization for the linear quadratic regulator to optimal control subject to hard safety constraints, and allows competing against a safety-aware clairvoyant policy with minor modifications. Numerical experiments confirm superior performance with respect to finite-horizon constrained $\mathcal{H}_2$ and $\mathcal{H}_\infty$ control laws when the disturbance realizations poorly fit classical assumptions.

📄 PDF Abstract BibTeX arXiv:2203.00358

Code (1)

decodepfl/safeminregret 공식 구현

Similar Papers 제목 키워드 기반

Safe Control of Partially-Observed Linear Time-Varying Systems with Minimal Worst-Case Dynamic Regret

2022-08-18 · HongYu Zhou, Vasileios Tzoumas

We present safe control of partially-observed linear time-varying systems in the presence of unknown and unpredictable process and measurement noise. We introduce a control algorithm that minimizes dynamic regret, i.e., …

Follow the Clairvoyant: an Imitation Learning Approach to Optimal Control

2022-11-14 · Andrea Martin, Luca Furieri, Florian Dörfler, John Lygeros 외

We consider control of dynamical systems through the lens of competitive analysis. Most prior work in this area focuses on minimizing regret, that is, the loss relative to an ideal clairvoyant policy that has noncausal a…

Imitation Learning

Foundations of Safe Online Reinforcement Learning in the Linear Quadratic Regulator: $\sqrt{T}$-Regret

2025-04-25 · Benjamin Schiffer, Lucas Janson

Understanding how to efficiently learn while adhering to safety constraints is essential for using online reinforcement learning in practical applications. However, proving rigorous regret bounds for safety-constrained r…

reinforcement-learningReinforcement Learning

Safety Filter for Robust Disturbance Rejection via Online Optimization

2024-11-14 · Joyce Lai, Peter Seiler

Disturbance rejection in high-precision control applications can be significantly improved upon via online convex optimization (OCO). This includes classical techniques such as recursive least squares (RLS) and more rece…

Safe Adaptive Learning-based Control for Constrained Linear Quadratic Regulators with Regret Guarantees

2021-10-31 · YingYing Li, Subhro Das, Jeff Shamma, Na Li

We study the adaptive control of an unknown linear system with a quadratic cost function subject to safety constraints on both the states and actions. The challenges of this problem arise from the tension among safety, e…