paper-with-me

홈 › Papers

Constrained Multi-Objective Reinforcement Learning with Max-Min Criterion

2026-05-29 · Giseung Park, Hyunyoung Nam, Woohyeon Byeon, Amir Leshem, Youngchul Sung arxiv

Multi-Objective Reinforcement Learning (MORL) extends standard RL by optimizing policies with respect to multiple, often conflicting, objectives. While max-min MORL has emerged as an effective approach for promoting fairness, its applicability remains limited, particularly when constraints must be incorporated. In this paper, we propose a MORL framework that integrates the max-min criterion with explicit constraint satisfaction. We establish a theoretical foundation for the proposed framework and validate the resulting algorithm through convergence analysis and experiments in tabular settings. We further demonstrate the practical relevance of our approach in simulated building thermal control, multi-objective locomotion control, and greenhouse-gas-emission-aware traffic management. Across these domains, our method effectively balances fairness and constraint satisfaction in multi-objective decision-making.

📄 PDF Abstract BibTeX arXiv:2605.31388

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning

Similar Papers 제목 키워드 기반

A Bayesian approach to constrained single- and multi-objective optimization

2015-10-02 · Paul Feliot, Julien Bect, Emmanuel Vazquez

This article addresses the problem of derivative-free (single- or multi-objective) optimization subject to multiple inequality constraints. Both the objective and constraint functions are assumed to be smooth, non-linear…

Multi-Objective Reinforcement Learning with Max-Min Criterion: A Game-Theoretic Approach

2025-10-23 · Woohyeon Byeon, Giseung Park, Jongseong Chae, Amir Leshem 외 arxiv

In this paper, we propose a provably convergent and practical framework for multi-objective reinforcement learning with max-min criterion. From a game-theoretic perspective, we reformulate max-min multi-objective reinfor…

Reinforcement Learning

Multi-Objective Reward and Preference Optimization: Theory and Algorithms

2025-12-11 · Akhil Agnihotri arxiv

This thesis develops theoretical frameworks and algorithms that advance constrained reinforcement learning (RL) across control, preference learning, and alignment of large language models. The first contribution addresse…

Reinforcement Learning

Expected Scalarised Returns Dominance: A New Solution Concept for Multi-Objective Decision Making

2021-06-02 · Conor F. Hayes, Timothy Verstraeten, Diederik M. Roijers, Enda Howley 외

In many real-world scenarios, the utility of a user is derived from the single execution of a policy. In this case, to apply multi-objective reinforcement learning, the expected utility of the returns must be optimised. …

Decision MakingMulti-Objective Reinforcement Learningreinforcement-learningReinforcement Learning+1

Finite-Time Analysis of Three-Timescale Constrained Actor-Critic and Constrained Natural Actor-Critic Algorithms

2023-10-25 · Prashansa Panda, Shalabh Bhatnagar

Actor Critic methods have found immense applications on a wide range of Reinforcement Learning tasks especially when the state-action space is large. In this paper, we consider actor critic and natural actor critic algor…