paper-with-me

홈 › Papers

Simplex-enabled Safe Continual Learning Machine

2024-09-05 · Hongpeng Cao, Yanbing Mao, Yihao Cai, Lui Sha, Marco Caccamo

This paper proposes the SeC-Learning Machine: Simplex-enabled safe continual learning for safety-critical autonomous systems. The SeC-learning machine is built on Simplex logic (that is, ``using simplicity to control complexity'') and physics-regulated deep reinforcement learning (Phy-DRL). The SeC-learning machine thus constitutes HP (high performance)-Student, HA (high assurance)-Teacher, and Coordinator. Specifically, the HP-Student is a pre-trained high-performance but not fully verified Phy-DRL, continuing to learn in a real plant to tune the action policy to be safe. In contrast, the HA-Teacher is a mission-reduced, physics-model-based, and verified design. As a complementary, HA-Teacher has two missions: backing up safety and correcting unsafe learning. The Coordinator triggers the interaction and the switch between HP-Student and HA-Teacher. Powered by the three interactive components, the SeC-learning machine can i) assure lifetime safety (i.e., safety guarantee in any continual-learning stage, regardless of HP-Student's success or convergence), ii) address the Sim2Real gap, and iii) learn to tolerate unknown unknowns in real plants. The experiments on a cart-pole system and a real quadruped robot demonstrate the distinguished features of the SeC-learning machine, compared with continual learning built on state-of-the-art safe DRL frameworks with approaches to addressing the Sim2Real gap.

📄 PDF Abstract BibTeX arXiv:2409.05898

Code (0)

등록된 구현이 없습니다.

Tasks

Continual LearningDeep Reinforcement Learning

Similar Papers 제목 키워드 기반

Dynamic-Weighted Simplex Strategy for Learning Enabled Cyber Physical Systems

2019-02-06 · Shreyas Ramakrishna, Charles Hartsell, Matthew P Burruss, Gabor Karsai 외

Cyber Physical Systems (CPS) have increasingly started using Learning Enabled Components (LECs) for performing perception-based control tasks. The simple design approach, and their capability to continuously learn has le…

Autonomous DrivingQ-LearningReinforcement Learning

Synergistic Simplex: Cooperative Runtime Assurance for Safety-Critical Autonomous Systems

2026-05-05 · Ayoosh Bansal, Mikael Yeghiazaryan, Artyom Khachatryan, Tianyi Zhu 외 arxiv

Autonomous systems increasingly rely on machine-learning (ML) components for safety-critical tasks such as perception and control in autonomous vehicles (AVs). While ML enables essential capabilities, it inevitably exhib…

Autonomous Vehicles

A Barrier Certificate-based Simplex Architecture for Systems with Approximate and Hybrid Dynamics

2022-02-20 · Amol Damare, Shouvik Roy, Roshan Sharma, Keith DSouza 외

We present Barrier-based Simplex (Bb-Simplex), a new, provably correct design for runtime assurance of continuous dynamical systems. Bb-Simplex is centered around the Simplex control architecture, which consists of a hig…

Dynamic Simplex: Balancing Safety and Performance in Autonomous Cyber Physical Systems

2023-02-20 · Baiting Luo, Shreyas Ramakrishna, Ava Pettet, Christopher Kuhn 외

Learning Enabled Components (LEC) have greatly assisted cyber-physical systems in achieving higher levels of autonomy. However, LEC's susceptibility to dynamic and uncertain operating conditions is a critical challenge f…

Decision MakingSequential Decision Making

SL1-Simplex: Safe Velocity Regulation of Self-Driving Vehicles in Dynamic and Unforeseen Environments

2020-08-04 · Yanbing Mao, Yuliang Gu, Naira Hovakimyan, Lui Sha 외

This paper proposes a novel extension of the Simplex architecture with model switching and model learning to achieve safe velocity regulation of self-driving vehicles in dynamic and unforeseen environments. To guarantee …

Autonomous Vehicles