paper-with-me

Papers

Optimal Belief Approximation

2016-10-27 · Reimar H. Leike, Torsten A. Enßlin

In Bayesian statistics probability distributions express beliefs. However, for many problems the beliefs cannot be computed analytically and approximations of beliefs are needed. We seek a loss function that quantifies how "embarrassing" it is to communicate a given approximation. We reproduce and discuss an old proof showing that there is only one ranking under the requirements that (1) the best ranked approximation is the non-approximated belief and (2) that the ranking judges approximations only by their predictions for actual outcomes. The loss function that is obtained in the derivation is equal to the Kullback-Leibler divergence when normalized. This loss function is frequently used in the literature. However, there seems to be confusion about the correct order in which its functional arguments, the approximated and non-approximated beliefs, should be used. The correct order ensures that the recipient of a communication is only deprived of the minimal amount of information. We hope that the elementary derivation settles the apparent confusion. For example when approximating beliefs with Gaussian distributions the optimal approximation is given by moment matching. This is in contrast to many suggested computational schemes.

📄 PDF Abstract BibTeX arXiv:1610.09018

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Finite Memory Belief Approximation for Optimal Control in Partially Observable Markov Decision Processes

2026-01-06 · Mintae Kim arxiv

We study finite memory belief approximation for partially observable (PO) stochastic optimal control (SOC) problems. While belief states are sufficient for SOC in partially observable Markov decision processes (POMDPs), …

Near Optimality of Finite Memory Feedback Policies in Partially Observed Markov Decision Processes

2020-10-15 · Ali Devran Kara, Serdar Yuksel

In the theory of Partially Observed Markov Decision Processes (POMDPs), existence of optimal policies have in general been established via converting the original partially observed stochastic control problem to a fully …

The Wasserstein Believer: Learning Belief Updates for Partially Observable Environments through Reliable Latent Space Models

2023-03-06 · Raphael Avalos, Florent Delgrange, Ann Nowé, Guillermo A. Pérez 외

Partially Observable Markov Decision Processes (POMDPs) are used to model environments where the full state cannot be perceived by an agent. As such the agent needs to reason taking into account the past observations and…

Neural Enhanced Belief Propagation on Factor Graphs

2020-03-04 · Victor Garcia Satorras, Max Welling

A graphical model is a structured representation of locally dependent random variables. A traditional method to reason over these random variables is to perform inference using belief propagation. When provided with the …

Under-Approximating Expected Total Rewards in POMDPs

2022-01-21 · Alexander Bork, Joost-Pieter Katoen, Tim Quatmann

We consider the problem: is the optimal expected total reward to reach a goal state in a partially observable Markov decision process (POMDP) below a given threshold? We tackle this -- generally undecidable -- problem by…