paper-with-me

Papers

Moody Learners -- Explaining Competitive Behaviour of Reinforcement Learning Agents

2020-07-30 · Pablo Barros, Ana Tanevska, Francisco Cruz, Alessandra Sciutti

Designing the decision-making processes of artificial agents that are involved in competitive interactions is a challenging task. In a competitive scenario, the agent does not only have a dynamic environment but also is directly affected by the opponents' actions. Observing the Q-values of the agent is usually a way of explaining its behavior, however, do not show the temporal-relation between the selected actions. We address this problem by proposing the \emph{Moody framework}. We evaluate our model by performing a series of experiments using the competitive multiplayer Chef's Hat card game and discuss how our model allows the agents' to obtain a holistic representation of the competitive dynamics within the game.

📄 PDF Abstract BibTeX arXiv:2007.16045

Code (0)

등록된 구현이 없습니다.

Tasks

Decision Makingreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

Loss aversion fosters coordination among independent reinforcement learners

2019-12-29 · Marco Jerome Gasparrini, Martí Sánchez-Fibla

We study what are the factors that can accelerate the emergence of collaborative behaviours among independent selfish learning agents. We depart from the "Battle of the Exes" (BoE), a spatial repeated game from which hum…

Reinforcement Learning

Explainability in autonomous pedagogically structured scenarios

2022-10-21 · Minal Suresh Patil

We present the notion of explainability for decision-making processes in a pedagogically structured autonomous environment. Multi-agent systems that are structured pedagogically consist of pedagogical teachers and learne…

Decision Making

Approximating Shapley Explanations in Reinforcement Learning

2025-11-08 · Daniel Beechey, Özgür Şimşek arxiv

Reinforcement learning has achieved remarkable success in complex decision-making environments, yet its lack of transparency limits its deployment in practice, especially in safety-critical settings. Shapley values from …

Reinforcement Learning

Overparameterization from Computational Constraints

2022-08-27 · Sanjam Garg, Somesh Jha, Saeed Mahloujifar, Mohammad Mahmoody 외

Overparameterized models with millions of parameters have been hugely successful. In this work, we ask: can the need for large models be, at least in part, due to the \emph{computational} limitations of the learner? Addi…

Mapping Monotonic Restrictions in Inductive Inference

2020-10-15 · Vanja Doskoč, Timo Kötzing

In language learning in the limit we investigate computable devices (learners) learning formal languages. Through the years, many natural restrictions have been imposed on the studied learners. As such, monotonic restric…