paper-with-me

홈 › Papers

Learned Lyapunov Shielding for Adaptive Control

2026-05-07 · Giansalvo Cirrincione, Adriano Fagiolini arxiv

We augment the Slotine--Li adaptive controller for Euler--Lagrange systems with three learned components: a structured-quadratic Lyapunov function \(V_ψ\) whose positive-definiteness follows from a Cholesky parameterization, a residual Soft Actor--Critic policy that adds bounded torque corrections to the analytic baseline, and a physics-informed neural network that estimates unmodeled dynamics. A closed-form safety filter, derived from the single affine constraint \(\dot V_ψ+ αV_ψ\le 0\), projects every policy output onto the safe set without requiring an online QP solver. We prove: global feasibility of the filter under a drift-decay condition on the control-degeneracy set; exponential stability under exact shielding, with a robust extension whose margin depends on the PINN approximation error; almost-sure convergence of the three-timescale policy--certificate--multiplier updates to a KKT point; and a PAC generalization bound for the certificate over compacts. On a 2-DOF manipulator with nonlinear friction and variable payload, the learned certificate accounts for most of the empirical gain: tracking error drops by 41\% on nominal friction and 24\% on aggressive friction at the centroid of the training distribution. A 7-DOF scalability study on a Franka Emika Panda confirms clean convergence of the full pipeline at industrial scale, identifies the conditions under which gains over exact model-based baselines should and should not be expected, and documents a warm-start pathology of the learned certificate that has practical implications for deployment.

📄 PDF Abstract BibTeX arXiv:2605.06934

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Follow the STARs: Dynamic $ω$-Regular Shielding of Learned Policies

2025-04-11 · Ashwani Anand, Satya Prakash Nayak, Ritam Raha, Anne-Kathrin Schmuck

This paper presents a novel dynamic post-shielding framework that enforces the full class of $\omega$-regular correctness properties over pre-computed probabilistic policies. This constitutes a paradigm shift from the pr…

MAMPS: Safe Multi-Agent Reinforcement Learning via Model Predictive Shielding

2019-10-25 · Wenbo Zhang, Osbert Bastani, Vijay Kumar

Reinforcement learning is a promising approach to learning control policies for performing complex multi-agent robotics tasks. However, a policy learned in simulation often fails to guarantee even simple safety propertie…

Multi-agent Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Lyapunov-Based Deep Residual Neural Network (ResNet) Adaptive Control

2024-04-10 · Omkar Sudhir Patil, Duc M. Le, Emily J. Griffis, Warren E. Dixon

Deep Neural Network (DNN)-based controllers have emerged as a tool to compensate for unstructured uncertainties in nonlinear dynamical systems. A recent breakthrough in the adaptive control literature provides a Lyapunov…

Adaptive GR(1) Specification Repair for Liveness-Preserving Shielding in Reinforcement Learning

2025-11-04 · Tiberiu-Andrei Georgescu, Alexander W. Goodall, Dalal Alrajeh, Francesco Belardinelli 외 arxiv

Shielding is widely used to enforce safety in reinforcement learning (RL), ensuring that an agent's actions remain compliant with formal specifications. Classical shielding approaches, however, are often static, in the s…

Inductive logic programmingReinforcement Learning

Transient and Asymptotic Properties of Robust Adaptive Controllers in the Presence of Non-Coercive Lyapunov Functions

2021-04-22 · Aditya A. Paranjape, Vivek Natarajan, Supratim Ghosh

Adaptive control architectures often make use of Lyapunov functions to design adaptive laws. We are specifically interested in adaptive control methods, such as the well-known L1 adaptive architecture, which employ a par…