paper-with-me

홈 › Papers

Ablation Study of How Run Time Assurance Impacts the Training and Performance of Reinforcement Learning Agents

2022-07-08 · Nathaniel Hamilton, Kyle Dunlap, Taylor T Johnson, Kerianne L Hobbs

Reinforcement Learning (RL) has become an increasingly important research area as the success of machine learning algorithms and methods grows. To combat the safety concerns surrounding the freedom given to RL agents while training, there has been an increase in work concerning Safe Reinforcement Learning (SRL). However, these new and safe methods have been held to less scrutiny than their unsafe counterparts. For instance, comparisons among safe methods often lack fair evaluation across similar initial condition bounds and hyperparameter settings, use poor evaluation metrics, and cherry-pick the best training runs rather than averaging over multiple random seeds. In this work, we conduct an ablation study using evaluation best practices to investigate the impact of run time assurance (RTA), which monitors the system state and intervenes to assure safety, on effective learning. By studying multiple RTA approaches in both on-policy and off-policy RL algorithms, we seek to understand which RTA methods are most effective, whether the agents become dependent on the RTA, and the importance of reward shaping versus safe exploration in RL agent training. Our conclusions shed light on the most promising directions of SRL, and our evaluation methodology lays the groundwork for creating better comparisons in future SRL work.

📄 PDF Abstract BibTeX arXiv:2207.04117

Code (0)

등록된 구현이 없습니다.

Tasks

reinforcement-learningReinforcement LearningReinforcement Learning (RL)Safe ExplorationSafe Reinforcement Learning

Similar Papers 제목 키워드 기반

AI Assurance: A Comprehensive Testing Strategy for Enterprise AI Systems

2026-05-22 · Chitra Badagi, Divye Singh, Animesh Sen, Adinath Shirsath arxiv

Enterprise AI systems, built on large language models, retrieval pipelines and autonomous agents, introduce a class of risks that traditional software quality assurance was never designed to address. These systems are pr…

DURA-CPS: A Multi-Role Orchestrator for Dependability Assurance in LLM-Enabled Cyber-Physical Systems

2025-06-04 · Trisanth Srinivasan, Santosh Patapati, Himani Musku, Idhant Gode 외

Cyber-Physical Systems (CPS) increasingly depend on advanced AI techniques to operate in critical applications. However, traditional verification and validation methods often struggle to handle the unpredictable and dyna…

Environmental Footprint of GenAI Research: Insights from the Moshi Foundation Model

2026-04-13 · Marta López-Rauhut, Loic Landrieu, Mathieu Aubry, Anne-Laure Ligozat arxiv

New multi-modal large language models (MLLMs) are continuously being trained and deployed, following rapid development cycles. This generative AI frenzy is driving steady increases in energy consumption, greenhouse gas e…

SOTER on ROS: A Run-Time Assurance Framework on the Robot Operating System

2020-08-21 · Sumukh Shivakumar, Hazem Torfah, Ankush Desai, Sanjit A. Seshia

We present an implementation of SOTER, a run-time assurance framework for building safe distributed mobile robotic (DMR) systems, on top of the Robot Operating System (ROS). The safety of DMR systems cannot always be gua…

Semi-Automated Quality Assurance in Digital Pathology: Tile Classification Approach

2025-06-12 · Meredith VandeHaar, M. Clinch, I. Yilmaz, M. A. Rahman 외

Quality assurance is a critical but underexplored area in digital pathology, where even minor artifacts can have significant effects. Artifacts have been shown to negatively impact the performance of AI diagnostic models…

Diagnosticwhole slide images