paper-with-me

Papers

Implicit Reparameterization Gradients

2018-05-22 · NeurIPS 2018 12 · Michael Figurnov, Shakir Mohamed, andriy mnih

By providing a simple and efficient way of computing low-variance gradients of continuous random variables, the reparameterization trick has become the technique of choice for training a variety of latent variable models. However, it is not applicable to a number of important continuous distributions. We introduce an alternative approach to computing reparameterization gradients based on implicit differentiation and demonstrate its broader applicability by applying it to Gamma, Beta, Dirichlet, and von Mises distributions, which cannot be used with the classic reparameterization trick. Our experiments show that the proposed approach is faster and more accurate than the existing gradient estimators for these distributions.

📄 PDF Abstract BibTeX arXiv:1805.08498

Code (2)

lucadellalib/sac-beta pytorch
sophieburkhardt/dirichlet-vae-topic-models tf

Similar Papers 제목 키워드 기반

Soft Actor-Critic with Beta Policy via Implicit Reparameterization Gradients

2024-09-08 · Luca Della Libera

Recent advances in deep reinforcement learning have achieved impressive results in a wide range of complex tasks, but poor sample efficiency remains a major obstacle to real-world deployment. Soft actor-critic (SAC) miti…

continuous-controlContinuous ControlDeep Reinforcement Learning

Dataset Distillation with Convexified Implicit Gradients

2023-02-13 · Noel Loo, Ramin Hasani, Mathias Lechner, Daniela Rus

We propose a new dataset distillation algorithm using reparameterization and convexification of implicit gradients (RCIG), that substantially improves the state-of-the-art. To this end, we first formulate dataset distill…

Dataset DistillationDataset Distillation - 1IPC

The Generalized Reparameterization Gradient

2016-10-07 · NeurIPS 2016 · Francisco J. R. Ruiz, Michalis K. Titsias, David M. Blei

The reparameterization gradient has become a widely used method to obtain Monte Carlo gradients to optimize the variational objective. However, this technique does not easily apply to commonly used distributions such as …

Variational Inference

The Generalized Reparameterization Gradient

2016-12-01 · NeurIPS 2016 12 · Francisco R. Ruiz, Michalis Titsias Rc Aueb, David Blei

The reparameterization gradient has become a widely used method to obtain Monte Carlo gradients to optimize the variational objective. However, this technique does not easily apply to commonly used distributions such as …

Variational Inference

Pathwise Derivatives Beyond the Reparameterization Trick

2018-06-05 · ICML 2018 7 · Martin Jankowiak, Fritz Obermeyer

We observe that gradients computed via the reparameterization trick are in direct correspondence with solutions of the transport equation in the formalism of optimal transport. We use this perspective to compute (approxi…

regressionVariational Inference