paper-with-me

Papers

An Actor-Critic Framework for Continuous-Time Jump-Diffusion Controls with Normalizing Flows

2026-04-07 · Liya Guo, Ruimeng Hu, Xu Yang, Yi Zhu arxiv

Continuous-time stochastic control with time-inhomogeneous jump-diffusion dynamics is central in finance and economics, but computing optimal policies is difficult under explicit time dependence, discontinuous shocks, and high dimensionality. We propose an actor-critic framework that serves as a mesh-free solver for entropy-regularized control problems and stochastic games with jumps. The approach is built on a time-inhomogeneous little q-function and an appropriate occupation measure, yielding a policy-gradient representation that accommodates time-dependent drift, volatility, and jump terms. To represent expressive stochastic policies in continuous-action spaces, we parameterize the actor using conditional normalizing flows, enabling flexible non-Gaussian policies while retaining exact likelihood evaluation for entropy regularization and policy optimization. We validate the method on time-inhomogeneous linear-quadratic control, Merton portfolio optimization, and a multi-agent portfolio game, using explicit solutions or high-accuracy benchmarks. Numerical results demonstrate stable learning under jump discontinuities, accurate approximation of optimal stochastic policies, and favorable scaling with respect to dimension and number of agents.

📄 PDF Abstract BibTeX arXiv:2604.05398

Code (0)

등록된 구현이 없습니다.

Tasks

Portfolio Optimization

Similar Papers 제목 키워드 기반

Reinforcement Learning for Jump-Diffusions, with Financial Applications

2024-05-26 · Xuefeng Gao, Lingfei Li, Xun Yu Zhou

We study continuous-time reinforcement learning (RL) for stochastic control in which system dynamics are governed by jump-diffusion processes. We formulate an entropy-regularized exploratory control problem with stochast…

Q-Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Continuous-time q-Learning for Jump-Diffusion Models under Tsallis Entropy

2024-07-04 · Lijun Bo, YiJie Huang, Xiang Yu, Tingting Zhang

This paper studies the continuous-time reinforcement learning in jump-diffusion models by featuring the q-learning (the continuous-time counterpart of Q-learning) under Tsallis entropy regularization. Contrary to the Sha…

Q-Learning

Learning a Unified Control Policy for Safe Falling

2017-03-08 · Visak CV Kumar, Sehoon Ha, C. Karen Liu

Being able to fall safely is a necessary motor skill for humanoids performing highly dynamic tasks, such as running and jumping. We propose a new method to learn a policy that minimizes the maximal impulse during the fal…

continuous-controlContinuous Control

Robust Reinforcement Learning under Diffusion Models for Data with Jumps

2024-11-18 · Chenyang Jiang, Donggyu Kim, Alejandra Quintos, Yazhen Wang

Reinforcement Learning (RL) has proven effective in solving complex decision-making tasks across various domains, but challenges remain in continuous-time settings, particularly when state dynamics are governed by stocha…

Decision Makingreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Exploratory Mean-Variance with Jumps: An Equilibrium Approach

2025-12-10 · Yuling Max Chen, Bin Li, David Saunders arxiv

Revisiting the continuous-time Mean-Variance (MV) Portfolio Optimization problem, we model the market dynamics with a jump-diffusion process and apply Reinforcement Learning (RL) techniques to facilitate informed explora…

Reinforcement LearningPortfolio Optimization