paper-with-me

홈 › Papers

Risk-Averse Reinforcement Learning: An Optimal Transport Perspective on Temporal Difference Learning

2025-02-22 · Zahra Shahrooei, Ali Baheri

The primary goal of reinforcement learning is to develop decision-making policies that prioritize optimal performance, frequently without considering risk or safety. In contrast, safe reinforcement learning seeks to reduce or avoid unsafe states. This letter introduces a risk-averse temporal difference algorithm that uses optimal transport theory to direct the agent toward predictable behavior. By incorporating a risk indicator, the agent learns to favor actions with predictable consequences. We evaluate the proposed algorithm in several case studies and show its effectiveness in the presence of uncertainty. The results demonstrate that our method reduces the frequency of visits to risky states while preserving performance. A Python implementation of the algorithm is available at https:// github.com/SAILRIT/Risk-averse-TD-Learning.

📄 PDF Abstract BibTeX arXiv:2502.16328

Code (1)

sailrit/risk-averse-td-learning 공식 구현

Tasks

Decision Makingreinforcement-learningReinforcement LearningSafe Reinforcement Learning

Similar Papers 제목 키워드 기반

Risk-Averse Learning by Temporal Difference Methods

2020-03-02 · Umit Kose, Andrzej Ruszczynski

We consider reinforcement learning with performance evaluated by a dynamic risk measure. We construct a projected risk-averse dynamic programming equation and study its properties. Then we propose risk-averse counterpart…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Optimal Transport and Risk Aversion in Kyle's Model of Informed Trading

2020-06-16 · Kerry Back, Francois Cocquemas, Ibrahim Ekren, Abraham Lioui

We establish connections between optimal transport theory and the dynamic version of the Kyle model, including new characterizations of informed trading profits via conjugate duality and Monge-Kantorovich duality. We use…

Risk-averse autonomous systems: A brief history and recent developments from the perspective of optimal control

2021-09-18 · Yuheng Wang, Margaret P. Chapman

We present an historical overview about the connections between the analysis of risk and the control of autonomous systems. We offer two main contributions. Our first contribution is to propose three overlapping paradigm…

RASR: Risk-Averse Soft-Robust MDPs with EVaR and Entropic Risk

2022-09-09 · Jia Lin Hau, Marek Petrik, Mohammad Ghavamzadeh, Reazul Russel

Prior work on safe Reinforcement Learning (RL) has studied risk-aversion to randomness in dynamics (aleatory) and to model uncertainty (epistemic) in isolation. We propose and analyze a new framework to jointly model the…

Reinforcement Learning (RL)Safe Reinforcement Learning

Finding Risk-Averse Shortest Path with Time-dependent Stochastic Costs

2017-01-03 · Dajian Li, Paul Weng, Orkun Karabasoglu

In this paper, we tackle the problem of risk-averse route planning in a transportation network with time-dependent and stochastic costs. To solve this problem, we propose an adaptation of the A* algorithm that accommodat…