paper-with-me

홈 › Papers

Entropy Regularization for Population Estimation

2022-08-24 · Ben Chugg, Peter Henderson, Jacob Goldin, Daniel E. Ho

Entropy regularization is known to improve exploration in sequential decision-making problems. We show that this same mechanism can also lead to nearly unbiased and lower-variance estimates of the mean reward in the optimize-and-estimate structured bandit setting. Mean reward estimation (i.e., population estimation) tasks have recently been shown to be essential for public policy settings where legal constraints often require precise estimates of population metrics. We show that leveraging entropy and KL divergence can yield a better trade-off between reward and estimator variance than existing baselines, all while remaining nearly unbiased. These properties of entropy regularization illustrate an exciting potential for bridging the optimal exploration and estimation literatures.

📄 PDF Abstract BibTeX arXiv:2208.11747

Code (0)

등록된 구현이 없습니다.

Tasks

Decision MakingSequential Decision Making

Methods 이 논문이 사용한 방법론

Entropy Regularization 설명 없음

Similar Papers 제목 키워드 기반

Exploratory LQG Mean Field Games with Entropy Regularization

2020-11-25 · Dena Firoozi, Sebastian Jaimungal

We study a general class of entropy-regularized multi-variate LQG mean field games (MFGs) in continuous time with $K$ distinct sub-population of agents. We extend the notion of actions to action distributions (explorator…

Towards Reliable WMH Segmentation under Domain Shift: An Application Study using Maximum Entropy Regularization to Improve Uncertainty Estimation

2025-06-17 · Franco Matzkin, Agostina Larrazabal, Diego H Milone, Jose Dolz 외

Accurate segmentation of white matter hyperintensities (WMH) is crucial for clinical decision-making, particularly in the context of multiple sclerosis. However, domain shifts, such as variations in MRI machine types or …

Decision MakingSegmentation

Promoting Stochasticity for Expressive Policies via a Simple and Efficient Regularization Method

2020-12-01 · NeurIPS 2020 12 · Qi Zhou, Yufei Kuang, Zherui Qiu, Houqiang Li 외

Many recent reinforcement learning (RL) methods learn stochastic policies with entropy regularization for exploration and robustness. However, in continuous action spaces, integrating entropy regularization with expressi…

continuous-controlContinuous Controlreinforcement-learningReinforcement Learning (RL)

Entropy-Regularized Partially Observed Markov Decision Processes

2021-12-22 · Timothy L. Molloy, Girish N. Nair

We investigate partially observed Markov decision processes (POMDPs) with cost functions regularized by entropy terms describing state, observation, and control uncertainty. Standard POMDP techniques are shown to offer b…

State Estimation

Tractable Regularization of Probabilistic Circuits

2021-06-04 · NeurIPS 2021 12 · Anji Liu, Guy Van Den Broeck

Probabilistic Circuits (PCs) are a promising avenue for probabilistic modeling. They combine advantages of probabilistic graphical models (PGMs) with those of neural networks (NNs). Crucially, however, they are tractable…

Density Estimation