paper-with-me

홈 › Papers

Risk-Sensitive Stochastic Optimal Control as Rao-Blackwellized Markovian Score Climbing

2023-12-21 · Hany Abdulsamad, Sahel Iqbal, Adrien Corenflos, Simo Särkkä

Stochastic optimal control of dynamical systems is a crucial challenge in sequential decision-making. Recently, control-as-inference approaches have had considerable success, providing a viable risk-sensitive framework to address the exploration-exploitation dilemma. Nonetheless, a majority of these techniques only invoke the inference-control duality to derive a modified risk objective that is then addressed within a reinforcement learning framework. This paper introduces a novel perspective by framing risk-sensitive stochastic control as Markovian score climbing under samples drawn from a conditional particle filter. Our approach, while purely inference-centric, provides asymptotically unbiased estimates for gradient-based policy optimization with optimal importance weighting and no explicit value function learning. To validate our methodology, we apply it to the task of learning neural non-Gaussian feedback policies, showcasing its efficacy on numerical benchmarks of stochastic dynamical systems.

📄 PDF Abstract BibTeX arXiv:2312.14000

Code (1)

hanyas/psoc 공식 구현 jax

Tasks

Decision MakingSequential Decision Making

Similar Papers 제목 키워드 기반

Risk-Sensitive Q-Learning in Continuous Time with Application to Dynamic Portfolio Selection

2025-12-02 · Chuhan Xie arxiv

This paper studies the problem of risk-sensitive reinforcement learning (RSRL) in continuous time, where the environment is characterized by a controllable stochastic differential equation (SDE) and the objective is a po…

Reinforcement Learning

A stochastic control perspective on term structure models with roll-over risk

2023-04-10 · Claudio Fontana, Simone Pavarana, Wolfgang J. Runggaldier

In this paper, we consider a generic interest rate market in the presence of roll-over risk, which generates spreads in spot/forward term rates. We do not require classical absence of arbitrage and rely instead on a mini…

Risk-sensitive control as inference with Rényi divergence

2024-11-04 · Kaito Ito, Kenji Kashima

This paper introduces the risk-sensitive control as inference (RCaI) that extends CaI by using R\'{e}nyi divergence variational inference. RCaI is shown to be equivalent to log-probability regularized risk-sensitive cont…

Reinforcement Learning (RL)Variational Inference

Distributionally Robust Path Integral Control

2023-10-02 · Hyuk Park, Duo Zhou, Grani A. Hanasusanto, Takashi Tanaka

We consider a continuous-time continuous-space stochastic optimal control problem, where the controller lacks exact knowledge of the underlying diffusion process, relying instead on a finite set of historical disturbance…

On the Global Convergence of Risk-Averse Policy Gradient Methods with Expected Conditional Risk Measures

2023-01-26 · Xian Yu, Lei Ying

Risk-sensitive reinforcement learning (RL) has become a popular tool for controlling the risk of uncertain outcomes and ensuring reliable performance in highly stochastic sequential decision-making problems. While Policy…

Decision MakingPolicy Gradient MethodsReinforcement Learning (RL)Sequential Decision Making