paper-with-me

홈 › Papers

Neural Actor-Critic Methods for Hamilton-Jacobi-Bellman PDEs: Asymptotic Analysis and Numerical Studies

2025-07-08 · Samuel N. Cohen, Jackson Hebner, Deqing Jiang, Justin Sirignano arxiv

We mathematically analyze and numerically study an actor-critic machine learning algorithm for solving high-dimensional Hamilton-Jacobi-Bellman (HJB) partial differential equations from stochastic control theory. The architecture of the critic (the estimator for the value function) is structured so that the boundary condition is always perfectly satisfied (rather than being included in the training loss) and utilizes a biased gradient which reduces computational cost. The actor (the estimator for the optimal control) is trained by minimizing the integral of the Hamiltonian over the domain, where the Hamiltonian is estimated using the critic. We show that the training dynamics of the actor and critic neural networks converge in a Sobolev-type space to a certain infinite-dimensional ordinary differential equation (ODE) as the number of hidden units in the actor and critic $\rightarrow \infty$. Further, under a convexity-like assumption on the Hamiltonian, we prove that any fixed point of this limit ODE is a solution of the original stochastic control problem. This provides an important guarantee for the algorithm's performance in light of the fact that finite-width neural networks may only converge to a local minimizers (and not optimal solutions) due to the non-convexity of their loss functions. In our numerical studies, we demonstrate that the algorithm can solve stochastic control problems accurately in up to 200 dimensions. In particular, we construct a series of increasingly complex stochastic control problems with known analytic solutions and study the algorithm's numerical performance on them. These problems range from a linear-quadratic regulator equation to highly challenging equations with non-convex Hamiltonians, allowing us to identify and analyze the strengths and limitations of this neural actor-critic method for solving HJB equations.

📄 PDF Abstract BibTeX arXiv:2507.06428

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Computation of Reachable Sets Based on Hamilton-Jacobi-Bellman Equation with Running Cost Function

2021-07-26 · Weiwei Liao, Tao Liang

A novel method for computing reachable sets is proposed in this paper. In the proposed method, a Hamilton-Jacobi-Bellman equation with running cost functionis numerically solved and the reachable sets of different time h…

On the Fragility of the Basis on the Hamilton-Jacobi-Bellman Equation in Economic Dynamics

2022-03-20 · Yuhki Hosoya

In this paper, we provide an example of the optimal growth model in which there exist infinitely many solutions to the Hamilton-Jacobi-Bellman equation but the value function does not satisfy this equation. We consider t…

DeepPAAC: A New Deep Galerkin Method for Principal-Agent Problems

2025-11-06 · Michael Ludkovski, Changgen Xie, Zimu Zhu arxiv

We consider numerical resolution of principal-agent (PA) problems in continuous time. We formulate a generic PA model with continuous and lump payments and a multi-dimensional strategy of the agent. To tackle the resulti…

Valuation of European Options under an Uncertain Market Price of Volatility Risk

2021-05-20 · Bartosz Jaroszkowski, Max Jensen

We propose a model to quantify the effect of parameter uncertainty on the option price in the Heston model. More precisely, we present a Hamilton-Jacobi-Bellman framework which allows us to evaluate best and worst case s…

Uncertainty Quantification

The Hamilton-Jacobi-Bellman Equation in Economic Dynamics with a Non-Smooth Fiscal Policy

2024-05-26 · Yuhki Hosoya

We consider a class of economic growth models that includes the classical Ramsey--Cass--Koopmans capital accumulation model and verify that, under several assumptions, the value function of the model is the unique viscos…