paper-with-me

Papers

Variational Inference MPC using Tsallis Divergence

2021-04-01 · Ziyi Wang, Oswin So, Jason Gibson, Bogdan Vlahov, Manan S. Gandhi, Guan-Horng Liu, Evangelos A. Theodorou

In this paper, we provide a generalized framework for Variational Inference-Stochastic Optimal Control by using thenon-extensive Tsallis divergence. By incorporating the deformed exponential function into the optimality likelihood function, a novel Tsallis Variational Inference-Model Predictive Control algorithm is derived, which includes prior works such as Variational Inference-Model Predictive Control, Model Predictive PathIntegral Control, Cross Entropy Method, and Stein VariationalInference Model Predictive Control as special cases. The proposed algorithm allows for effective control of the cost/reward transform and is characterized by superior performance in terms of mean and variance reduction of the associated cost. The aforementioned features are supported by a theoretical and numerical analysis on the level of risk sensitivity of the proposed algorithm as well as simulation experiments on 5 different robotic systems with 3 different policy parameterizations.

📄 PDF Abstract BibTeX arXiv:2104.00241

Code (0)

등록된 구현이 없습니다.

Tasks

Model Predictive ControlVariational Inference

Similar Papers 제목 키워드 기반

General Munchausen Reinforcement Learning with Tsallis Kullback-Leibler Divergence

2023-09-21 · NeurIPS 2023 11

Many policy optimization approaches in reinforcement learning incorporate a Kullback-Leilbler (KL) divergence to the previous policy, to prevent the policy from changing too quickly. This idea was initially proposed in a…

On Divergence Measures for Training GFlowNets

2024-10-12 · Tiago da Silva, Eliezer de Souza da Silva, Diego Mesquita

Generative Flow Networks (GFlowNets) are amortized inference models designed to sample from unnormalized distributions over composable objects, with applications in generative modeling for tasks in fields such as causal …

Causal DiscoveryDrug DiscoveryVariational Inference

Generalized Munchausen Reinforcement Learning using Tsallis KL Divergence

2023-01-27 · Lingwei Zhu, Zheng Chen, Matthew Schlegel, Martha White

Many policy optimization approaches in reinforcement learning incorporate a Kullback-Leilbler (KL) divergence to the previous policy, to prevent the policy from changing too quickly. This idea was initially proposed in a…

Atari Gamesreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Calibeating for general proper losses: A Bregman divergence approach

2026-05-17 · Maximilian Fichtl, Cristóbal Guzmán, Nishant A. Mehta arxiv

This work introduces a general framework for calibeating based on regret minimization. As compared to Foster and Hart's seminal calibeating work which had specialized treatments of Brier score (squared loss) and log loss…

On Voronoi diagrams and dual Delaunay complexes on the information-geometric Cauchy manifolds

2020-06-12 · Frank Nielsen

We study the Voronoi diagrams of a finite set of Cauchy distributions and their dual complexes from the viewpoint of information geometry by considering the Fisher-Rao distance, the Kullback-Leibler divergence, the chi s…