paper-with-me

Papers

Policy Gradient with Expected Quadratic Utility Maximization: A New Mean-Variance Approach in Reinforcement Learning

2020-09-28 · Masahiro Kato, Kei Nakagawa

In real-world decision-making problems, risk management is critical. Among various risk management approaches, the mean-variance criterion is one of the most widely used in practice. In this paper, we suggest expected quadratic utility maximization (EQUM) as a new framework for policy gradient style reinforcement learning (RL) algorithms with mean-variance control. The quadratic utility function is a common objective of risk management in finance and economics. The proposed EQUM framework has several interpretations, such as reward-constrained variance minimization and regularization, as well as agent utility maximization. In addition, the computation of the EQUM framework is easier than that of existing mean-variance RL methods, which require double sampling. In experiments, we demonstrate the effectiveness of the proposed framework in benchmark setting of RL and financial data.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Decision MakingManagementReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

Optimal investment and consumption under logarithmic utility and uncertainty model

2022-11-10 · Wahid Faidi

We study a robust utility maximization problem in the case of an incomplete market and logarithmic utility with general stochastic constraints, not necessarily convex. Our problem is equivalent to maximizing of nonlinear…

Learning-NUM: Network Utility Maximization with Unknown Utility Functions and Queueing Delay

2020-12-16 · Xinzhe Fu, Eytan Modiano

Network Utility Maximization (NUM) studies the problems of allocating traffic rates to network users in order to maximize the users' total utility subject to network resource constraints. In this paper, we propose a new …

Scheduling

Optimal investment and consumption under $g$- expected utility and general constraints in incomplete market

2025-01-27 · Wahid Faidi

This article studies the problem of utility maximization in an incomplete market under a class of nonlinear expectations and general constraints on trading strategies. Using a $g$-martingale method, we provide an explici…

Utility maximization under endogenous pricing

2020-05-08 · Thai Nguyen, Mitja Stadje

We study the expected utility maximization problem of a large investor who is allowed to make transactions on tradable assets in an incomplete financial market with endogenous permanent market impacts. The asset prices a…

Convergence and sample complexity of natural policy gradient primal-dual methods for constrained MDPs

2022-06-06 · Dongsheng Ding, Kaiqing Zhang, Jiali Duan, Tamer Başar 외

We study sequential decision making problems aimed at maximizing the expected total reward while satisfying a constraint on the expected total utility. We employ the natural policy gradient method to solve the discounted…

Decision MakingSequential Decision Making