paper-with-me

홈 › Papers

Learning Exceptionality and Variation with Lexically Scaled MaxEnt

2019-01-01 · WS 2019 1 · Coral Hughto, Andrew Lamont, Br Prickett, on, Gaja Jarosz
📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Joint learning of constraint weights and gradient inputs in Gradient Symbolic Computation with constrained optimization

2020-07-01 · WS 2020 7 · Max Nelson

This paper proposes a method for the joint optimization of constraint weights and symbol activations within the Gradient Symbolic Computation (GSC) framework. The set of grammars representable in GSC is proven to be a su…

S$^2$AC: Energy-Based Reinforcement Learning with Stein Soft Actor Critic

2024-05-02 · Safa Messaoud, Billel Mokeddem, Zhenghai Xue, Linsey Pang 외

Learning expressive stochastic policies instead of deterministic ones has been proposed to achieve better stability, sample complexity, and robustness. Notably, in Maximum Entropy Reinforcement Learning (MaxEnt RL), the …

MuJoCoVariational Inference

Local Exceptionality Detection in Time Series Using Subgroup Discovery

2021-08-05 · Dan Hudson, Travis J. Wiltshire, Martin Atzmueller

In this paper, we present a novel approach for local exceptionality detection on time series data. This method provides the ability to discover interpretable patterns in the data, which can be used to understand and pred…

Subgroup DiscoveryTime SeriesTime Series Analysis

Risk-sensitive control as inference with Rényi divergence

2024-11-04 · Kaito Ito, Kenji Kashima

This paper introduces the risk-sensitive control as inference (RCaI) that extends CaI by using R\'{e}nyi divergence variational inference. RCaI is shown to be equivalent to log-probability regularized risk-sensitive cont…

Reinforcement Learning (RL)Variational Inference

If MaxEnt RL is the Answer, What is the Question?

2019-10-04 · Benjamin Eysenbach, Sergey Levine

Experimentally, it has been observed that humans and animals often make decisions that do not maximize their expected utility, but rather choose outcomes randomly, with probability proportional to expected utility. Proba…

Reinforcement Learning