paper-with-me

홈 › Papers

Expected Scalarised Returns Dominance: A New Solution Concept for Multi-Objective Decision Making

2021-06-02 · Conor F. Hayes, Timothy Verstraeten, Diederik M. Roijers, Enda Howley, Patrick Mannion

In many real-world scenarios, the utility of a user is derived from the single execution of a policy. In this case, to apply multi-objective reinforcement learning, the expected utility of the returns must be optimised. Various scenarios exist where a user's preferences over objectives (also known as the utility function) are unknown or difficult to specify. In such scenarios, a set of optimal policies must be learned. However, settings where the expected utility must be maximised have been largely overlooked by the multi-objective reinforcement learning community and, as a consequence, a set of optimal solutions has yet to be defined. In this paper we address this challenge by proposing first-order stochastic dominance as a criterion to build solution sets to maximise expected utility. We also propose a new dominance criterion, known as expected scalarised returns (ESR) dominance, that extends first-order stochastic dominance to allow a set of optimal policies to be learned in practice. We then define a new solution concept called the ESR set, which is a set of policies that are ESR dominant. Finally, we define a new multi-objective distributional tabular reinforcement learning (MOT-DRL) algorithm to learn the ESR set in a multi-objective multi-armed bandit setting.

📄 PDF Abstract BibTeX arXiv:2106.01048

Code (0)

등록된 구현이 없습니다.

Tasks

Decision MakingMulti-Objective Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

Multi-Objective Multi-Agent Decision Making: A Utility-based Analysis and Survey

2019-09-06 · Roxana Rădulescu, Patrick Mannion, Diederik M. Roijers, Ann Nowé

The majority of multi-agent system (MAS) implementations aim to optimise agents' policies with respect to a single objective, despite the fact that many real-world problem domains are inherently multi-objective in nature…

Decision Making

A utility-based analysis of equilibria in multi-objective normal form games

2020-01-17 · Roxana Rădulescu, Patrick Mannion, Yijie Zhang, Diederik M. Roijers 외

In multi-objective multi-agent systems (MOMAS), agents explicitly consider the possible tradeoffs between conflicting objective functions. We argue that compromises between competing objectives in MOMAS should be analyse…

Form

Multi-Objective Coordination Graphs for the Expected Scalarised Returns with Generative Flow Models

2022-07-01 · Conor F. Hayes, Timothy Verstraeten, Diederik M. Roijers, Enda Howley 외

Many real-world problems contain multiple objectives and agents, where a trade-off exists between objectives. Key to solving such problems is to exploit sparse dependency structures that exist between agents. For example…

Multi-Objective Reinforcement Learning

Loss Aversion and State-Dependent Linear Utility Functions for Monetary Returns

2024-10-24 · Somdeb Lahiri

We present a theory of expected utility with state-dependent linear utility functions for monetary returns, that incorporates the possibility of loss-aversion. Our results relate to first order stochastic dominance, mean…

PySDTest: a Python/Stata Package for Stochastic Dominance Tests

2023-07-20 · Kyungho Lee, Yoon-Jae Whang

We introduce PySDTest, a Python/Stata package for statistical tests of stochastic dominance. PySDTest implements various testing procedures such as Barrett and Donald (2003), Linton et al. (2005), Linton et al. (2010), a…