Training Characteristic Functions with Reinforcement Learning: XAI-methods play Connect Four
One of the goals of Explainable AI (XAI) is to determine which input components were relevant for a classifier decision. This is commonly know as saliency attribution. Characteristic functions (from cooperative game theory) are able to evaluate partial inputs and form the basis for theoretically "fair" attribution methods like Shapley values. Given only a standard classifier function, it is unclear how partial input should be realised. Instead, most XAI-methods for black-box classifiers like neural networks consider counterfactual inputs that generally lie off-manifold. This makes them hard to evaluate and easy to manipulate. We propose a setup to directly train characteristic functions in the form of neural networks to play simple two-player games. We apply this to the game of Connect Four by randomly hiding colour information from our agents during training. This has three advantages for comparing XAI-methods: It alleviates the ambiguity about how to realise partial input, makes off-manifold evaluation unnecessary and allows us to compare the methods by letting them play against each other.
Code (0)
등록된 구현이 없습니다.
Tasks
counterfactualExplainable Artificial Intelligence (XAI)reinforcement-learningReinforcement Learning (RL)Similar Papers 제목 키워드 기반
Playing Atari with Capsule Networks: A systematic comparison of CNN and CapsNets-based agents.
In recent years, Capsule Networks (CapsNets) have achieved promising results in tasks in the object recognition task thanks to their invariance characteristics towards pose and lighting. They have been proposed as an alt…
Deep Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)Task Independent Capsule-Based Agents for Deep Q-Learning
In recent years, Capsule Networks (CapsNets) have achieved promising results in tasks such as object recognition thanks to their invariance characteristics towards pose and lighting. They have been proposed as an alterna…
Deep Reinforcement LearningObject RecognitionQ-Learningreinforcement-learning+2Efficient Reinforcement Learning by Reducing Forgetting with Elephant Activation Functions
Catastrophic forgetting has remained a significant challenge for efficient reinforcement learning for decades (Ring 1994, Rivest and Precup 2003). While recent works have proposed effective methods to mitigate this issue…
Reinforcement LearningTransformable Gaussian Reward Function for Socially-Aware Navigation with Deep Reinforcement Learning
Robot navigation has transitioned from prioritizing obstacle avoidance to adopting socially aware navigation strategies that accommodate human presence. As a result, the recognition of socially aware navigation within dy…
Deep Reinforcement Learningreinforcement-learningReinforcement LearningRobot NavigationLearning to Play Two-Player Perfect-Information Games without Knowledge
In this paper, several techniques for learning game state evaluation functions by reinforcement are proposed. The first is a generalization of tree bootstrapping (tree learning): it is adapted to the context of reinforce…
reinforcement-learningReinforcement LearningReinforcement Learning (RL)Vocal Bursts Valence Prediction