paper-with-me

홈 › Papers

Gradient Flows for Regularized Stochastic Control Problems

2020-06-10 · David Šiška, Łukasz Szpruch

This paper studies stochastic control problems with the action space taken to be probability measures, with the objective penalised by the relative entropy. We identify suitable metric space on which we construct a gradient flow for the measure-valued control process, in the set of admissible controls, along which the cost functional is guaranteed to decrease. It is shown that any invariant measure of this gradient flow satisfies the Pontryagin optimality principle. If the problem we work with is sufficiently convex, the gradient flow converges exponentially fast. Furthermore, the optimal measure-valued control process admits a Bayesian interpretation which means that one can incorporate prior knowledge when solving such stochastic control problems. This work is motivated by a desire to extend the theoretical underpinning for the convergence of stochastic gradient type algorithms widely employed in the reinforcement learning community to solve control problems.

📄 PDF Abstract BibTeX arXiv:2006.05956

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

A Distributed Algorithm for Measure-valued Optimization with Additive Objective

2022-02-17 · Iman Nodozi, Abhishek Halder

We propose a distributed nonparametric algorithm for solving measure-valued optimization problems with additive objectives. Such problems arise in several contexts in stochastic learning and control including Langevin sa…

An Actor-Critic Framework for Continuous-Time Jump-Diffusion Controls with Normalizing Flows

2026-04-07 · Liya Guo, Ruimeng Hu, Xu Yang, Yi Zhu arxiv

Continuous-time stochastic control with time-inhomogeneous jump-diffusion dynamics is central in finance and economics, but computing optimal policies is difficult under explicit time dependence, discontinuous shocks, an…

Portfolio Optimization

Global Convergence of Policy Gradient for Entropy Regularized Linear-Quadratic Control with Multiplicative Noise

2025-10-03 · Gabriel Diaz, Lucky Li, Wenhao Zhang arxiv

Reinforcement Learning (RL) has emerged as a powerful framework for sequential decision-making in dynamic environments, particularly when system parameters are unknown. This paper investigates RL-based control for entrop…

Reinforcement Learning

Controlling the Flow: Stability and Convergence for Stochastic Gradient Descent with Decaying Regularization

2025-05-16 · Sebastian Kassing, Simon Weissmann, Leif Döring

The present article studies the minimization of convex, L-smooth functions defined on a separable real Hilbert space. We analyze regularized stochastic gradient descent (reg-SGD), a variant of stochastic gradient descent…

Image Reconstruction

RES: Regularized Stochastic BFGS Algorithm

2014-01-29 · Aryan Mokhtari, Alejandro Ribeiro

RES, a regularized stochastic version of the Broyden-Fletcher-Goldfarb-Shanno (BFGS) quasi-Newton method is proposed to solve convex optimization problems with stochastic objectives. The use of stochastic gradient descen…

Second-order methods