paper-with-me

Papers

A Kernel Perspective on Behavioural Metrics for Markov Decision Processes

2023-10-05 · Pablo Samuel Castro, Tyler Kastner, Prakash Panangaden, Mark Rowland

Behavioural metrics have been shown to be an effective mechanism for constructing representations in reinforcement learning. We present a novel perspective on behavioural metrics for Markov decision processes via the use of positive definite kernels. We leverage this new perspective to define a new metric that is provably equivalent to the recently introduced MICo distance (Castro et al., 2021). The kernel perspective further enables us to provide new theoretical results, which has so far eluded prior work. These include bounding value function differences by means of our metric, and the demonstration that our metric can be provably embedded into a finite-dimensional Euclidean space with low distortion error. These are two crucial properties when using behavioural metrics for reinforcement learning representations. We complement our theory with strong empirical results that demonstrate the effectiveness of these methods in practice.

📄 PDF Abstract BibTeX arXiv:2310.19804

Code (0)

등록된 구현이 없습니다.

Tasks

reinforcement-learningReinforcement Learning

Similar Papers 제목 키워드 기반

The behavioural aspect of green technology investments: a general positive model in the context of heterogeneous agents

2016-03-22

Studies report that firms do not invest in cost-effective green technologies. While economic barriers can explain parts of the gap, behavioural aspects cause further under-valuation. This could be partly due to systemati…

Decision MakingDiversity

Heterogeneous Hidden Markov Models for Sleep Activity Recognition from Multi-Source Passively Sensed Data

2022-11-08 · Fernando Moreno-Pino, María Martínez-García, Pablo M. Olmos, Antonio Artés-Rodríguez

Psychiatric patients' passive activity monitoring is crucial to detect behavioural shifts in real-time, comprising a tool that helps clinicians supervise patients' evolution over time and enhance the associated treatment…

Activity Recognition

On Imperfect Recall in Multi-Agent Influence Diagrams

2023-07-11 · James Fox, Matt MacDermott, Lewis Hammond, Paul Harrenstein 외

Multi-agent influence diagrams (MAIDs) are a popular game-theoretic model based on Bayesian networks. In some settings, MAIDs offer significant advantages over extensive-form game representations. Previous work on MAIDs …

A Meta-Bayesian Model of Intentional Visual Search

2020-06-05 · Maell Cullen, Jonathan Monney, M. Berk Mirza, Rosalyn Moran

We propose a computational model of visual search that incorporates Bayesian interpretations of the neural mechanisms that underlie categorical perception and saccade planning. To enable meaningful comparisons between si…

Decision Making

MICo: Improved representations via sampling-based state similarity for Markov decision processes

2021-06-03 · NeurIPS 2021 12 · Pablo Samuel Castro, Tyler Kastner, Prakash Panangaden, Mark Rowland

We present a new behavioural distance over the state space of a Markov decision process, and demonstrate the use of this distance as an effective means of shaping the learnt representations of deep reinforcement learning…

Atari GamesDeep Reinforcement Learningreinforcement-learningReinforcement Learning (RL)