paper-with-me

홈 › Papers

Rate-Regularization and Generalization in VAEs

2019-11-11 · Alican Bozkurt, Babak Esmaeili, Jean-Baptiste Tristan, Dana H. Brooks, Jennifer G. Dy, Jan-Willem van de Meent

Variational autoencoders optimize an objective that combines a reconstruction loss (the distortion) and a KL term (the rate). The rate is an upper bound on the mutual information, which is often interpreted as a regularizer that controls the degree of compression. We here examine whether inclusion of the rate also acts as an inductive bias that improves generalization. We perform rate-distortion analyses that control the strength of the rate term, the network capacity, and the difficulty of the generalization problem. Decreasing the strength of the rate paradoxically improves generalization in most settings, and reducing the mutual information typically leads to underfitting. Moreover, we show that generalization continues to improve even after the mutual information saturates, indicating that the gap on the bound (i.e. the KL divergence relative to the inference marginal) affects generalization. This suggests that the standard Gaussian prior is not an inductive bias that typically aids generalization, prompting work to understand what choices of priors improve generalization in VAEs.

📄 PDF Abstract BibTeX arXiv:1911.04594

Code (0)

등록된 구현이 없습니다.

Tasks

Inductive Bias

Methods 이 논문이 사용한 방법론

USD Coin Customer Service Number +1-833-534-1729 설명 없음

Similar Papers 제목 키워드 기반

Information-theoretic Generalization Analysis for VQ-VAEs: A Role of Latent Variables

2025-05-26 · Futoshi Futami, Masahiro Fujisawa

Latent variables (LVs) play a crucial role in encoder-decoder models by enabling effective data compression, prediction, and generation. Although their theoretical properties, such as generalization, have been extensivel…

Data CompressionDecoder

Amortized Inference Regularization

2018-05-23 · NeurIPS 2018 12 · Rui Shu, Hung H. Bui, Shengjia Zhao, Mykel J. Kochenderfer 외

The variational autoencoder (VAE) is a popular model for density estimation and representation learning. Canonically, the variational principle suggests to prefer an expressive inference model so that the variational app…

Density EstimationRepresentation Learning

High-dimensional Asymptotics of VAEs: Threshold of Posterior Collapse and Dataset-Size Dependence of Rate-Distortion Curve

2023-09-14 · Yuma Ichikawa, Koji Hukushima

In variational autoencoders (VAEs), the variational posterior often collapses to the prior, known as posterior collapse, which leads to poor representation learning quality. An adjustable hyperparameter beta has been int…

Representation Learning

Consistency Regularization for Variational Auto-Encoders

2021-05-31 · NeurIPS 2021 12 · Samarth Sinha, Adji B. Dieng

Variational auto-encoders (VAEs) are a powerful approach to unsupervised learning. They enable scalable approximate posterior inference in latent-variable models using variational inference (VI). A VAE posits a variation…

Image GenerationVariational Inference

A Bayesian Nonparametrics View into Deep Representations

2020-12-01 · NeurIPS 2020 12 · Michał Jamroż, Marcin Kurdziel, Mateusz Opala

We investigate neural network representations from a probabilistic perspective. Specifically, we leverage Bayesian nonparametrics to construct models of neural activations in Convolutional Neural Networks (CNNs) and late…

Memorization