paper-with-me

Papers

Exploratory Control with Tsallis Entropy for Latent Factor Models

2022-11-14 · Ryan Donnelly, Sebastian Jaimungal

We study optimal control in models with latent factors where the agent controls the distribution over actions, rather than actions themselves, in both discrete and continuous time. To encourage exploration of the state space, we reward exploration with Tsallis Entropy and derive the optimal distribution over states - which we prove is $q$-Gaussian distributed with location characterized through the solution of an FBS$\Delta$E and FBSDE in discrete and continuous time, respectively. We discuss the relation between the solutions of the optimal exploration problems and the standard dynamic optimal control solution. Finally, we develop the optimal policy in a model-agnostic setting along the lines of soft $Q$-learning. The approach may be applied in, e.g., developing more robust statistical arbitrage trading strategies.

📄 PDF Abstract BibTeX arXiv:2211.07622

Code (0)

등록된 구현이 없습니다.

Tasks

Q-Learning

Similar Papers 제목 키워드 기반

Exploratory Utility Maximization Problem with Tsallis Entropy

2025-02-03 · Chen Ziyi, Gu Jia-wen

We study expected utility maximization problem with constant relative risk aversion utility function in a complete market under the reinforcement learning framework. To induce exploration, we introduce the Tsallis entrop…

reinforcement-learningReinforcement Learning

Tsallis Reinforcement Learning: A Unified Framework for Maximum Entropy Reinforcement Learning

2019-01-31 · Kyungjae Lee, Sungyub Kim, Sungbin Lim, Sungjoon Choi 외

In this paper, we present a new class of Markov decision processes (MDPs), called Tsallis MDPs, with Tsallis entropy maximization, which generalizes existing maximum entropy reinforcement learning (RL). A Tsallis MDP pro…

MuJoCoreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Continuous-time q-Learning for Jump-Diffusion Models under Tsallis Entropy

2024-07-04 · Lijun Bo, YiJie Huang, Xiang Yu, Tingting Zhang

This paper studies the continuous-time reinforcement learning in jump-diffusion models by featuring the q-learning (the continuous-time counterpart of Q-learning) under Tsallis entropy regularization. Contrary to the Sha…

Q-Learning

Tsallis Entropy Regularization for Linearly Solvable MDP and Linear Quadratic Regulator

2024-03-04 · Yota Hashizume, Koshi Oishi, Kenji Kashima

Shannon entropy regularization is widely adopted in optimal control due to its ability to promote exploration and enhance robustness, e.g., maximum entropy reinforcement learning known as Soft Actor-Critic. In this paper…

reinforcement-learningReinforcement Learning

Unifying Decision Trees Split Criteria Using Tsallis Entropy

2015-11-25 · Yisen Wang, Chaobing Song, Shu-Tao Xia

The construction of efficient and effective decision trees remains a key topic in machine learning because of their simplicity and flexibility. A lot of heuristic algorithms have been proposed to construct near-optimal d…