Discrete-Time Mean-Variance Strategy Based on Reinforcement Learning
This paper studies a discrete-time mean-variance model based on reinforcement learning. Compared with its continuous-time counterpart in \cite{zhou2020mv}, the discrete-time model makes more general assumptions about the asset's return distribution. Using entropy to measure the cost of exploration, we derive the optimal investment strategy, whose density function is also Gaussian type. Additionally, we design the corresponding reinforcement learning algorithm. Both simulation experiments and empirical analysis indicate that our discrete-time model exhibits better applicability when analyzing real-world data than the continuous-time model.
Code (0)
등록된 구현이 없습니다.
Tasks
reinforcement-learningReinforcement LearningSimilar Papers 제목 키워드 기반
Discrete time multi-period mean-variance model: Bellman type strategy and Empirical analysis
In this paper, we attempt to introduce the Bellman principle for a discrete time multi-period mean-variance model. Based on this new take on the Bellman principle, we obtain a dynamic time-consistent optimal strategy and…
Reinforcement Learning for a Discrete-Time Linear-Quadratic Control Problem with an Application
We study the discrete-time linear-quadratic (LQ) control model using reinforcement learning (RL). Using entropy to measure the cost of exploration, we prove that the optimal feedback policy for the problem must be Gaussi…
ManagementReinforcement Learning (RL)Backpropagation through the Void: Optimizing control variates for black-box gradient estimation
Gradient-based optimization is the foundation of deep learning and reinforcement learning. Even when the mechanism being optimized is unknown or not differentiable, optimization using high-variance or biased gradient est…
reinforcement-learningReinforcement LearningReinforcement Learning (RL)Time Consistent Stopping For The Mean-Standard Deviation Problem --- The Discrete Time Case
Inspired by Strotz's consistent planning strategy, we formulate the infinite horizon mean-variance stopping problem as a subgame perfect Nash equilibrium in order to determine time consistent strategies with no regret. E…
Evolutionary game with stochastic payoffs in a finite island model
In this paper, we consider a two-player two-strategy game with random payoffs in a population subdivided into $d$ demes, each containing $N$ individuals at the beginning of any given generation and experiencing local ext…