paper-with-me

Papers

Decoupling Time and Risk: Risk-Sensitive Reinforcement Learning with General Discounting

2026-02-04 · Mehrdad Moghimi, Anthony Coache, Hyejin Ku arxiv

Distributional reinforcement learning (RL) is a powerful framework increasingly adopted in safety-critical domains for its ability to optimize risk-sensitive objectives. However, the role of the discount factor is often overlooked, as it is typically treated as a fixed parameter of the Markov decision process or tunable hyperparameter, with little consideration of its effect on the learned policy. In the literature, it is well-known that the discounting function plays a major role in characterizing time preferences of an agent, which an exponential discount factor cannot fully capture. Building on this insight, we propose a novel framework that supports flexible discounting of future rewards and optimization of risk measures in distributional RL. We provide a technical analysis of the optimality of our algorithms, show that our multi-horizon extension fixes issues raised with existing methodologies, and validate the robustness of our methods through extensive experiments. Our results highlight that discounting is a cornerstone in decision-making problems for capturing more expressive temporal and risk preferences profiles, with potential implications for real-world safety-critical applications.

📄 PDF Abstract BibTeX arXiv:2602.04131

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning

Similar Papers 제목 키워드 기반

Risk-Sensitive Reinforcement Learning: Near-Optimal Risk-Sample Tradeoff in Regret

2020-06-22 · NeurIPS 2020 12 · Yingjie Fei, Zhuoran Yang, Yudong Chen, Zhaoran Wang 외

We study risk-sensitive reinforcement learning in episodic Markov decision processes with unknown transition kernels, where the goal is to optimize the total reward under the risk measure of exponential utility. We propo…

Q-Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Distributional Model Equivalence for Risk-Sensitive Reinforcement Learning

2023-07-04 · NeurIPS 2023 11 · Tyler Kastner, Murat A. Erdogdu, Amir-Massoud Farahmand

We consider the problem of learning models for risk-sensitive reinforcement learning. We theoretically demonstrate that proper value equivalence, a method of learning models which can be used to plan optimally in the ris…

Distributional Reinforcement Learningmodelreinforcement-learningReinforcement Learning

Continuous-time Risk-sensitive Reinforcement Learning via Quadratic Variation Penalty

2024-04-19 · Yanwei Jia

This paper studies continuous-time risk-sensitive reinforcement learning (RL) under the entropy-regularized, exploratory diffusion process formulation with the exponential-form objective. The risk-sensitive objective ari…

Q-Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Risk-sensitive reinforcement learning using expectiles, shortfall risk and optimized certainty equivalent risk

2026-02-10 · Sumedh Gupte, Shrey Rakeshkumar Patel, Soumen Pachal, Prashanth L. A. 외 arxiv

We propose risk-sensitive reinforcement learning algorithms catering to three families of risk measures, namely expectiles, utility-based shortfall risk and optimized certainty equivalent risk. For each risk measure, in …

Reinforcement Learning

Is Risk-Sensitive Reinforcement Learning Properly Resolved?

2023-07-02 · Ruiwen Zhou, Minghuan Liu, Kan Ren, Xufang Luo 외

Due to the nature of risk management in learning applicable policies, risk-sensitive reinforcement learning (RSRL) has been realized as an important direction. RSRL is usually achieved by learning risk-sensitive objectiv…

Distributional Reinforcement LearningManagementQ-Learningreinforcement-learning+1