paper-with-me

Papers

Discrete-Time Mean-Variance Strategy Based on Reinforcement Learning

2023-12-24 · Xiangyu Cui, Xun Li, Yun Shi, Si Zhao

This paper studies a discrete-time mean-variance model based on reinforcement learning. Compared with its continuous-time counterpart in \cite{zhou2020mv}, the discrete-time model makes more general assumptions about the asset's return distribution. Using entropy to measure the cost of exploration, we derive the optimal investment strategy, whose density function is also Gaussian type. Additionally, we design the corresponding reinforcement learning algorithm. Both simulation experiments and empirical analysis indicate that our discrete-time model exhibits better applicability when analyzing real-world data than the continuous-time model.

📄 PDF Abstract BibTeX arXiv:2312.15385

Code (0)

등록된 구현이 없습니다.

Tasks

reinforcement-learningReinforcement Learning

Similar Papers 제목 키워드 기반

Discrete time multi-period mean-variance model: Bellman type strategy and Empirical analysis

2020-11-22 · Shuzhen Yang

In this paper, we attempt to introduce the Bellman principle for a discrete time multi-period mean-variance model. Based on this new take on the Bellman principle, we obtain a dynamic time-consistent optimal strategy and…

Reinforcement Learning for a Discrete-Time Linear-Quadratic Control Problem with an Application

2024-12-08 · Lucky Li

We study the discrete-time linear-quadratic (LQ) control model using reinforcement learning (RL). Using entropy to measure the cost of exploration, we prove that the optimal feedback policy for the problem must be Gaussi…

ManagementReinforcement Learning (RL)

Backpropagation through the Void: Optimizing control variates for black-box gradient estimation

2017-10-31 · ICLR 2018 1 · Will Grathwohl, Dami Choi, Yuhuai Wu, Geoffrey Roeder 외

Gradient-based optimization is the foundation of deep learning and reinforcement learning. Even when the mechanism being optimized is unknown or not differentiable, optimization using high-variance or biased gradient est…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Time Consistent Stopping For The Mean-Standard Deviation Problem --- The Discrete Time Case

2019-04-19

Inspired by Strotz's consistent planning strategy, we formulate the infinite horizon mean-variance stopping problem as a subgame perfect Nash equilibrium in order to determine time consistent strategies with no regret. E…

Evolutionary game with stochastic payoffs in a finite island model

2023-11-01 · Dhaker Kroumi, Sabin Lessard

In this paper, we consider a two-player two-strategy game with random payoffs in a population subdivided into $d$ demes, each containing $N$ individuals at the beginning of any given generation and experiencing local ext…