paper-with-me

홈 › Papers

Trading Information between Latents in Hierarchical Variational Autoencoders

2023-02-09 · Tim Z. Xiao, Robert Bamler

Variational Autoencoders (VAEs) were originally motivated (Kingma & Welling, 2014) as probabilistic generative models in which one performs approximate Bayesian inference. The proposal of $\beta$-VAEs (Higgins et al., 2017) breaks this interpretation and generalizes VAEs to application domains beyond generative modeling (e.g., representation learning, clustering, or lossy data compression) by introducing an objective function that allows practitioners to trade off between the information content ("bit rate") of the latent representation and the distortion of reconstructed data (Alemi et al., 2018). In this paper, we reconsider this rate/distortion trade-off in the context of hierarchical VAEs, i.e., VAEs with more than one layer of latent variables. We identify a general class of inference models for which one can split the rate into contributions from each layer, which can then be tuned independently. We derive theoretical bounds on the performance of downstream tasks as functions of the individual layers' rates and verify our theoretical findings in large-scale experiments. Our results provide guidance for practitioners on which region in rate-space to target for a given application.

📄 PDF Abstract BibTeX arXiv:2302.04855

Code (1)

timxzz/hit 공식 구현 pytorch

Tasks

Bayesian InferenceData CompressionRepresentation Learning

Similar Papers 제목 키워드 기반

Mutual Information Constraints for Monte-Carlo Objectives

2020-12-01 · Gábor Melis, András György, Phil Blunsom

A common failure mode of density models trained as variational autoencoders is to model the data without relying on their latent variables, rendering these variables useless. Two contributing factors, the underspecificat…

Effective Use of Variational Embedding Capacity in Expressive End-to-End Speech Synthesis

2019-06-08 · Eric Battenberg, Soroosh Mariooryad, Daisy Stanton, RJ Skerry-Ryan 외

Recent work has explored sequence-to-sequence latent variable models for expressive speech synthesis (supporting control and transfer of prosody and style), but has not presented a coherent framework for understanding th…

Expressive Speech SynthesisSpeech SynthesisStyle Transfer

GFlowNet-EM for learning compositional latent variable models

2023-02-13 · Edward J. Hu, Nikolay Malkin, Moksh Jain, Katie Everett 외

Latent variable models (LVMs) with discrete compositional latents are an important but challenging setting due to a combinatorially large number of possible configurations of the latents. A key tradeoff in modeling the p…

Variational Inference

Learned Image Compression with Hierarchical Progressive Context Modeling

2025-07-25 · Yuqi Li, Haotian Zhang, Li Li, Dong Liu arxiv

Context modeling is essential in learned image compression for accurately estimating the distribution of latents. While recent advanced methods have expanded context modeling capacity, they still struggle to efficiently …

Image Compression

A Cross Channel Context Model for Latents in Deep Image Compression

2021-03-04 · Changyue Ma, Zhao Wang, Ruling Liao, Yan Ye

This paper presents a cross channel context model for latents in deep image compression. Generally, deep image compression is based on an autoencoder framework, which transforms the original image to latents at the encod…

Image CompressionMS-SSIMSSIM