paper-with-me

홈 › Papers

Multi-Objective Reinforcement Learning with Max-Min Criterion: A Game-Theoretic Approach

2025-10-23 · Woohyeon Byeon, Giseung Park, Jongseong Chae, Amir Leshem, Youngchul Sung arxiv

In this paper, we propose a provably convergent and practical framework for multi-objective reinforcement learning with max-min criterion. From a game-theoretic perspective, we reformulate max-min multi-objective reinforcement learning as a two-player zero-sum regularized continuous game and introduce an efficient algorithm based on mirror descent. Our approach simplifies the policy update while ensuring global last-iterate convergence. We provide a comprehensive theoretical analysis on our algorithm, including iteration complexity under both exact and approximate policy evaluations, as well as sample complexity bounds. To further enhance performance, we modify the proposed algorithm with adaptive regularization. Our experiments demonstrate the convergence behavior of the proposed algorithm in tabular settings, and our implementation for deep reinforcement learning significantly outperforms previous baselines in many MORL environments.

📄 PDF Abstract BibTeX arXiv:2510.20235

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning

Similar Papers 제목 키워드 기반

Fairness in Multi-Agent Sequential Decision-Making

2014-12-01 · NeurIPS 2014 12 · Chongjie Zhang, Julie A. Shah

We define a fairness solution criterion for multi-agent decision-making problems, where agents have local interests. This new criterion aims to maximize the worst performance of agents with consideration on the overall p…

Decision MakingFairnessSequential Decision Making

Constrained Multi-Objective Reinforcement Learning with Max-Min Criterion

2026-05-29 · Giseung Park, Hyunyoung Nam, Woohyeon Byeon, Amir Leshem 외 arxiv

Multi-Objective Reinforcement Learning (MORL) extends standard RL by optimizing policies with respect to multiple, often conflicting, objectives. While max-min MORL has emerged as an effective approach for promoting fair…

Reinforcement Learning

Continuous time mean-variance-utility portfolio problem and its equilibrium strategy

2020-05-14 · Ben-Zhang Yang, Xin-Jiang He, Song-Ping Zhu

In this paper, we propose a new class of optimization problems, which maximize the terminal wealth and accumulated consumption utility subject to a mean variance criterion controlling the final risk of the portfolio. The…

GARL: Game-Theoretic Reinforcement Learning for Multi-Agent Strategic Prioritisation

2026-06-03 · Yuxiao Ye, Yiwen Zhang, Huiyuan Xie, Yuqin Huang 외 arxiv

LLM-based multi-agent systems are increasingly used for strategic decision-making tasks. In such settings, performance depends not only on individual model capabilities, but also on the policies by which agents interact …

Multi-agent Reinforcement Learning

Competitive Multi-agent Inverse Reinforcement Learning with Sub-optimal Demonstrations

2018-01-07 · ICML 2018 7 · Xingyu Wang, Diego Klabjan

This paper considers the problem of inverse reinforcement learning in zero-sum stochastic games when expert demonstrations are known to be not optimal. Compared to previous works that decouple agents in the game by assum…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)