paper-with-me

Papers

Risk-Aware General-Utility Markov Decision Processes

2026-07-10 · Pedro P. Santos, Fábio Vital, Alberto Sardinha, Francisco S. Melo arxiv

We study general-utility Markov decision processes (GUMDPs) with risk-aware objectives. In this framework, an agent aims to optimize a risk measure of the distribution of objective values, where the objective function depends on the frequency of visitation of states induced by the agent's policy. First, we motivate, propose, and formalize risk-aware GUMDPs, which enable agents and decision makers to trade off expected performance by risk aversion while benefiting from the rich set of objectives that can be cast under the framework of GUMDPs. We focus our attention on the entropic risk measure (ERM). Second, we show how we can solve risk-aware GUMDPs with ERM objectives by resorting to online planning techniques. In particular, we propose an approach based on Monte Carlo Tree Search (MCTS) to provably solve risk-aware GUMDPs up to any desired accuracy. Third, we provide a set of experimental results showcasing that our approach is successful when optimizing for a spectrum of risk-aware behaviors in the context of GUMDPs under diverse tasks (standard MDPs, maximum state entropy exploration, imitation learning, and multi-objective MDPs).

📄 PDF Abstract BibTeX arXiv:2607.09298

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Planning and Learning in Average Risk-aware MDPs

2025-03-22 · Weikai Wang, Erick Delage

For continuing tasks, average cost Markov decision processes have well-documented value and can be solved using efficient algorithms. However, it explicitly assumes that the agent is risk-neutral. In this work, we extend…

Q-Learning

Risk-sensitive Markov Decision Process and Learning under General Utility Functions

2023-11-22 · Zhengqi Wu, Renyuan Xu

Reinforcement Learning (RL) has gained substantial attention across diverse application domains and theoretical investigations. Existing literature on RL theory largely focuses on risk-neutral settings where the decision…

Reinforcement Learning (RL)

Learning Utilities from Demonstrations in Markov Decision Processes

2024-09-25 · Filippo Lazzati, Alberto Maria Metelli

Our goal is to extract useful knowledge from demonstrations of behavior in sequential decision-making problems. Although it is well-known that humans commonly engage in risk-sensitive behaviors in the presence of stochas…

Decision MakingSequential Decision Making

Beyond Accuracy: A Decision-Theoretic Framework for Allocation-Aware Healthcare AI

2026-01-06 · Rifa Ferzana arxiv

Artificial intelligence (AI) systems increasingly achieve expert-level predictive accuracy in healthcare, yet improvements in model performance often fail to produce corresponding gains in patient outcomes. We term this …

TRUST-ESD: A Risk-Calibrated and Governance-Aware AI Framework for Enterprise Strategic Decision Support Under Uncertainty

2026-07-22 · Tian Qiu, Li Yan, Mahabubur Rahman Miraj, Shanqin Yi 외 arxiv

Enterprise strategic decision support requires AI systems that are not only accurate, but also uncertainty-aware, risk-calibrated, explainable, and governance-compliant. This paper proposes TRUST-ESD, a risk-calibrated a…