Imitation Learning with Stability and Safety Guarantees
A method is presented to learn neural network (NN) controllers with stability and safety guarantees through imitation learning (IL). Convex stability and safety conditions are derived for linear time-invariant plant dynamics with NN controllers by merging Lyapunov theory with local quadratic constraints to bound the nonlinear activation functions in the NN. These conditions are incorporated in the IL process, which minimizes the IL loss, and maximizes the volume of the region of attraction associated with the NN controller simultaneously. An alternating direction method of multipliers based algorithm is proposed to solve the IL problem. The method is illustrated on an inverted pendulum system, aircraft longitudinal dynamics, and vehicle lateral dynamics.
Code (1)
Tasks
Imitation LearningSimilar Papers 제목 키워드 기반
MPC as a Copilot: A Predictive Filter Framework with Safety and Stability Guarantees
Ensuring both safety and stability remains a fundamental challenge in learning-based control, where goal-oriented policies often neglect system constraints and closed-loop state convergence. To address this limitation, t…
Optimal Sequencing and Motion Control in a Roundabout with Safety Guarantees
This paper develops a controller for Connected and Automated Vehicles (CAVs) traversing a single-lane roundabout. The controller simultaneously determines the optimal sequence and associated optimal motion control jointl…
Model Predictive ControlReinforcement Learning for Distributed Transient Frequency Control with Stability and Safety Guarantees
This paper proposes a reinforcement learning-based approach for optimal transient frequency control in power systems with stability and safety guarantees. Building on Lyapunov stability theory and safety-critical control…
reinforcement-learningReinforcement LearningReinforcement Learning (RL)Adaptive Estimation-Based Safety-Critical Cruise Control of Vehicular Platoons
Optimal cruise control design can increase highway throughput and vehicle safety in traffic flow. In most heterogeneous platoons, the absence of vehicle-to-vehicle (V2V) communication poses challenges in maintaining syst…
Globally Stable Neural Imitation Policies
Imitation learning presents an effective approach to alleviate the resource-intensive and time-consuming nature of policy learning from scratch in the solution space. Even though the resulting policy can mimic expert dem…
Imitation Learning