Choquet regularization for reinforcement learning
We propose \emph{Choquet regularizers} to measure and manage the level of exploration for reinforcement learning (RL), and reformulate the continuous-time entropy-regularized RL problem of Wang et al. (2020, JMLR, 21(198)) in which we replace the differential entropy used for regularization with a Choquet regularizer. We derive the Hamilton--Jacobi--Bellman equation of the problem, and solve it explicitly in the linear--quadratic (LQ) case via maximizing statically a mean--variance constrained Choquet regularizer. Under the LQ setting, we derive explicit optimal distributions for several specific Choquet regularizers, and conversely identify the Choquet regularizers that generate a number of broadly used exploratory samplers such as $\epsilon$-greedy, exponential, uniform and Gaussian.
Code (0)
등록된 구현이 없습니다.
Tasks
reinforcement-learningReinforcement LearningReinforcement Learning (RL)Similar Papers 제목 키워드 기반
Quantiles under ambiguity and risk sharing
Choquet capacities and integrals are central concepts in decision making under ambiguity or model uncertainty, pioneered by Schmeidler. Motivated by risk optimization problems for quantiles under ambiguity, we study the …
Decision MakingIdentification of Choquet capacity in multicriteria sorting problems through stochastic inverse analysis
In multicriteria decision aiding (MCDA), the Choquet integral has been used as an aggregation operator to deal with the case of interacting decision criteria. While the application of the Choquet integral for ranking pro…
DescriptiveChoquet rating criteria, risk measures, and risk consistency
Credit ratings are widely used by investors as a screening device. We introduce and study several natural notions of risk consistency that promote prudent investment decisions in the framework of Choquet rating criteria.…
Using the Choquet Integral in the Pooling Layer in Deep Learning Networks
This paper aims to introduce the proposal of replacing the usual pooling functions by the Choquet integral in Deep Learning Networks. The Choquet integral is an aggregation function studied and applied in several areas, …
Fuzzy Rough Choquet Distances for Classification
This paper introduces a novel Choquet distance using fuzzy rough set based measures. The proposed distance measure combines the attribute information received from fuzzy rough set theory with the flexibility of the Choqu…
AttributeClassification