paper-with-me

Papers

Utilizing Explainability Techniques for Reinforcement Learning Model Assurance

2023-11-27 · Alexander Tapley, Kyle Gatesman, Luis Robaina, Brett Bissey, Joseph Weissman

Explainable Reinforcement Learning (XRL) can provide transparency into the decision-making process of a Deep Reinforcement Learning (DRL) model and increase user trust and adoption in real-world use cases. By utilizing XRL techniques, researchers can identify potential vulnerabilities within a trained DRL model prior to deployment, therefore limiting the potential for mission failure or mistakes by the system. This paper introduces the ARLIN (Assured RL Model Interrogation) Toolkit, an open-source Python library that identifies potential vulnerabilities and critical points within trained DRL models through detailed, human-interpretable explainability outputs. To illustrate ARLIN's effectiveness, we provide explainability visualizations and vulnerability analysis for a publicly available DRL model. The open-source code repository is available for download at https://github.com/mitre/arlin.

📄 PDF Abstract BibTeX arXiv:2311.15838

Code (1)

mitre/arlin 공식 구현

Tasks

Decision MakingDeep Reinforcement Learningmodelreinforcement-learningReinforcement Learning

Methods 이 논문이 사용한 방법론

Library 설명 없음

Similar Papers 제목 키워드 기반

PCB Component Detection using Computer Vision for Hardware Assurance

2022-02-17 · Wenwei Zhao, Suprith Gurudu, Shayan Taheri, Shajib Ghosh 외

Printed Circuit Board (PCB) assurance in the optical domain is a crucial field of study. Though there are many existing PCB assurance methods using image processing, computer vision (CV), and machine learning (ML), the P…

Embedding Explainable AI in NHS Clinical Safety: The Explainability-Enabled Clinical Safety Framework (ECSF)

2025-10-24 · Robert Gigiu arxiv

Artificial intelligence (AI) is increasingly embedded in NHS workflows, but its probabilistic and adaptive behaviour conflicts with the deterministic assumptions underpinning existing clinical-safety standards. DCB0129 a…

Runway Sign Classifier: A DAL C Certifiable Machine Learning System

2023-10-10 · Konstantin Dmitriev, Johann Schumann, Islam Bostanov, Mostafa Abdelhamid 외

In recent years, the remarkable progress of Machine Learning (ML) technologies within the domain of Artificial Intelligence (AI) systems has presented unprecedented opportunities for the aviation industry, paving the way…

Management

The Role of Explainability in Assuring Safety of Machine Learning in Healthcare

2021-09-01 · Yan Jia, John McDermid, Tom Lawton, Ibrahim Habli

Established approaches to assuring safety-critical systems and software are difficult to apply to systems employing ML where there is no clear, pre-defined specification against which to assess validity. This problem is …

BIG-bench Machine LearningExplainable Artificial Intelligence (XAI)

Evaluating Explainability in Safety-Critical ATR Systems: Limitations of Post-Hoc Methods and Paths Toward Robust XAI

2026-05-07 · Vanessa Buhrmester, David Muench, Dimitri Bulatov, Michael Arens arxiv

Explainable Artificial Intelligence (XAI) is increasingly rec ognized as essential for deploying machine learning systems in safety critical environments. In Automatic Target Recognition (ATR), where models operate on im…

Decision Making