paper-with-me

Papers

Counterfactual Explanation Policies in RL

2023-07-25 · Shripad V. Deshmukh, Srivatsan R, Supriti Vijay, Jayakumar Subramanian, Chirag Agarwal

As Reinforcement Learning (RL) agents are increasingly employed in diverse decision-making problems using reward preferences, it becomes important to ensure that policies learned by these frameworks in mapping observations to a probability distribution of the possible actions are explainable. However, there is little to no work in the systematic understanding of these complex policies in a contrastive manner, i.e., what minimal changes to the policy would improve/worsen its performance to a desired level. In this work, we present COUNTERPOL, the first framework to analyze RL policies using counterfactual explanations in the form of minimal changes to the policy that lead to the desired outcome. We do so by incorporating counterfactuals in supervised learning in RL with the target outcome regulated using desired return. We establish a theoretical connection between Counterpol and widely used trust region-based policy optimization methods in RL. Extensive empirical analysis shows the efficacy of COUNTERPOL in generating explanations for (un)learning skills while keeping close to the original policy. Our results on five different RL environments with diverse state and action spaces demonstrate the utility of counterfactual explanations, paving the way for new frontiers in designing and developing counterfactual policies.

📄 PDF Abstract BibTeX arXiv:2307.13192

Code (0)

등록된 구현이 없습니다.

Tasks

counterfactualCounterfactual ExplanationDecision MakingReinforcement Learning (RL)

Methods 이 논문이 사용한 방법론

Counterfactuals 설명 없음

Similar Papers 제목 키워드 기반

Decisions, Counterfactual Explanations and Strategic Behavior

2020-02-11 · NeurIPS 2020 12 · Stratis Tsirtsis, Manuel Gomez-Rodriguez

As data-driven predictive models are increasingly used to inform decisions, it has been argued that decision makers should provide explanations that help individuals understand what would have to change for these decisio…

counterfactual

Leveraging Counterfactual Paths for Contrastive Explanations of POMDP Policies

2024-03-28 · Benjamin Kraske, Zakariya Laouar, Zachary Sunberg

As humans come to rely on autonomous systems more, ensuring the transparency of such systems is important to their continued adoption. Explainable Artificial Intelligence (XAI) aims to reduce confusion and foster trust i…

counterfactualExplainable artificial intelligenceExplainable Artificial Intelligence (XAI)

Disagreement amongst counterfactual explanations: How transparency can be deceptive

2023-04-25 · Dieter Brughmans, Lissa Melis, David Martens

Counterfactual explanations are increasingly used as an Explainable Artificial Intelligence (XAI) technique to provide stakeholders of complex machine learning algorithms with explanations for data-driven decisions. The …

counterfactualDecision MakingDiversityExplainable artificial intelligence+1

TalkToAgent: A Human-centric Explanation of Reinforcement Learning Agents with Large Language Models

2025-09-05 · Haechang Kim, Hao Chen, Can Li, Jong Min Lee arxiv

Explainable Reinforcement Learning (XRL) has emerged as a promising approach in improving the transparency of Reinforcement Learning (RL) agents. However, there remains a gap between complex RL policies and domain expert…

Reinforcement Learning

ACTER: Diverse and Actionable Counterfactual Sequences for Explaining and Diagnosing RL Policies

2024-02-09 · Jasmina Gajcin, Ivana Dusparic

Understanding how failure occurs and how it can be prevented in reinforcement learning (RL) is necessary to enable debugging, maintain user trust, and develop personalized policies. Counterfactual reasoning has often bee…

counterfactualCounterfactual ReasoningDiversityreinforcement-learning+1