paper-with-me

Papers

Physics-model-guided Worst-case Sampling for Safe Reinforcement Learning

2024-12-17 · Hongpeng Cao, Yanbing Mao, Lui Sha, Marco Caccamo

Real-world accidents in learning-enabled CPS frequently occur in challenging corner cases. During the training of deep reinforcement learning (DRL) policy, the standard setup for training conditions is either fixed at a single initial condition or uniformly sampled from the admissible state space. This setup often overlooks the challenging but safety-critical corner cases. To bridge this gap, this paper proposes a physics-model-guided worst-case sampling strategy for training safe policies that can handle safety-critical cases toward guaranteed safety. Furthermore, we integrate the proposed worst-case sampling strategy into the physics-regulated deep reinforcement learning (Phy-DRL) framework to build a more data-efficient and safe learning algorithm for safety-critical CPS. We validate the proposed training strategy with Phy-DRL through extensive experiments on a simulated cart-pole system, a 2D quadrotor, a simulated and a real quadruped robot, showing remarkably improved sampling efficiency to learn more robust safe policies.

📄 PDF Abstract BibTeX arXiv:2412.13224

Code (0)

등록된 구현이 없습니다.

Tasks

Deep Reinforcement Learningreinforcement-learningReinforcement LearningSafe Reinforcement Learning

Similar Papers 제목 키워드 기반

Physics-Informed Neural Networks for Minimising Worst-Case Violations in DC Optimal Power Flow

2021-06-28 · Rahul Nellikkath, Spyros Chatzivasileiadis

Physics-informed neural networks exploit the existing models of the underlying physical systems to generate higher accuracy results with fewer data. Such approaches can help drastically reduce the computation time and ge…

Safety-Aware Imitation Learning via MPC-Guided Disturbance Injection

2025-08-05 · Le Qiu, Yusuf Umut Ciftci, Somil Bansal arxiv

Imitation Learning has provided a promising approach to learning complex robot behaviors from expert demonstrations. However, learned policies can make errors that lead to safety violations, which limits their deployment…

MADR: MPC-guided Adversarial DeepReach

2025-10-21 · Ryan Teoh, Sander Tonkens, William Sharpless, Aijia Yang 외 arxiv

Hamilton-Jacobi (HJ) Reachability offers a framework for generating safe value functions and policies in the face of adversarial disturbance, but is limited by the curse of dimensionality. Physics-informed deep learning …

Self-Supervised Learning

SafeFlow: Real-Time Text-Driven Humanoid Whole-Body Control via Physics-Guided Rectified Flow and Selective Safety Gating

2026-03-25 · Hanbyel Cho, Sang-Hun Kim, Jeonguk Kang, Donghan Koo arxiv

Recent advances in real-time interactive text-driven motion generation have enabled humanoids to perform diverse behaviors. However, kinematics-only generators often exhibit physical hallucinations, producing motion traj…

Learning Density Distribution of Reachable States for Autonomous Systems

2021-09-14 · Yue Meng, Dawei Sun, Zeng Qiu, Md Tawhid Bin Waez 외

State density distribution, in contrast to worst-case reachability, can be leveraged for safety-related problems to better quantify the likelihood of the risk for potentially hazardous situations. In this work, we propos…