paper-with-me

홈 › Papers

SHIELD: Scalable Optimal Control with Certification using Duality and Convexity

2026-05-09 · Hansung Kim, Siddharth H. Nair, Francesco Borrelli arxiv

We present SHIELD, a hierarchical algorithm that reduces both the decision-variable dimension and the constraint set in $\ell_1$-regularized convex programs. From strong convexity and Lagrangian duality, we derive certificates that \emph{safely} discard constraints and decision variables while guaranteeing that all removed constraints remain satisfied and all removed variables are null. To further accelerate the proposed algorithm, we propose a transformer-based deep neural network to guide the dual certificate inference. We validate SHIELD on stochastic model predictive control (SMPC) in complex, multi-modal traffic scenarios, comparing against a full-dimensional SMPC policy. Numerical simulations demonstrate order-of-magnitude computational speedups while preserving feasibility and closed-loop safety, highlighting the practicality of certifiably safe, lightweight MPC in complex driving scenes.

📄 PDF Abstract BibTeX arXiv:2605.09171

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Shielded Reinforcement Learning for Hybrid Systems

2023-08-28 · Asger Horn Brorholt, Peter Gjøl Jensen, Kim Guldstrand Larsen, Florian Lorber 외

Safe and optimal controller synthesis for switched-controlled hybrid systems, which combine differential equations and discrete changes of the system's state, is known to be intricately hard. Reinforcement learning has b…

reinforcement-learningReinforcement Learning

ISAACS: Iterative Soft Adversarial Actor-Critic for Safety

2022-12-06 · Kai-Chieh Hsu, Duy Phuong Nguyen, Jaime Fernández Fisac

The deployment of robots in uncontrolled environments requires them to operate robustly under previously unseen scenarios, like irregular terrain and wind conditions. Unfortunately, while rigorous safety frameworks from …

Shielded Analysis: Certification and Characterization of Defensibility in Systems under Adversarial Interaction

2026-06-11 · Achraf Hsain, Sultan Almuhammadi arxiv

Formal safety analysis determines whether a system admits a safe defense; adaptive evaluation characterizes the operating quality sustained under adversarial interaction. Both answers matter because systems with the same…

Multi-agent Reinforcement Learning

An efficient nonconvex reformulation of stagewise convex optimization problems

2020-10-27 · NeurIPS 2020 12 · Rudy Bunel, Oliver Hinder, Srinadh Bhojanapalli, Krishnamurthy 외

Convex optimization problems with staged structure appear in several contexts, including optimal control, verification of deep neural networks, and isotonic regression. Off-the-shelf solvers can solve these problems but …

Dual Formulation of the Optimal Consumption problem with Multiplicative Habit Formation

2025-02-19 · Thijs Kamma, Antoon Pelsser

This paper provides a dual formulation of the optimal consumption problem with internal multiplicative habit formation. In this problem, the agent derives utility from the ratio of consumption to the internal habit compo…