Bi-level Doubly Variational Learning for Energy-based Latent Variable Models
Energy-based latent variable models (EBLVMs) are more expressive than conventional energy-based models. However, its potential on visual tasks are limited by its training process based on maximum likelihood estimate that requires sampling from two intractable distributions. In this paper, we propose Bi-level doubly variational learning (BiDVL), which is based on a new bi-level optimization framework and two tractable variational distributions to facilitate learning EBLVMs. Particularly, we lead a decoupled EBLVM consisting of a marginal energy-based distribution and a structural posterior to handle the difficulties when learning deep EBLVMs on images. By choosing a symmetric KL divergence in the lower level of our framework, a compact BiDVL for visual tasks can be obtained. Our model achieves impressive image generation performance over related works. It also demonstrates the significant capacity of testing image reconstruction and out-of-distribution detection.
Code (0)
등록된 구현이 없습니다.
Tasks
Image GenerationImage ReconstructionOut-of-Distribution DetectionSimilar Papers 제목 키워드 기반
Doubly Stochastic Variational Inference for Neural Processes with Hierarchical Latent Variables
Neural processes (NPs) constitute a family of variational approximate models for stochastic processes with promising properties in computational efficiency and uncertainty quantification. These processes use neural netwo…
Computational EfficiencyregressionUncertainty QuantificationVariational InferenceAutomatic Relevance Determination For Deep Generative Models
A recurring problem when building probabilistic latent variable models is regularization and model selection, for instance, the choice of the dimensionality of the latent space. In the context of belief networks with lat…
Model SelectionVariational InferenceGeneralised Gaussian Process Latent Variable Models (GPLVM) with Stochastic Variational Inference
Gaussian process latent variable models (GPLVM) are a flexible and non-linear approach to dimensionality reduction, extending classical Gaussian processes to an unsupervised learning context. The Bayesian incarnation of …
BenchmarkingDimensionality ReductionGaussian ProcessesVariational InferenceBi-level Score Matching for Learning Energy-based Latent Variable Models
Score matching (SM) provides a compelling approach to learn energy-based models (EBMs) by avoiding the calculation of partition function. However, it remains largely open to learn energy-based latent variable models (EBL…
Rolling Shutter CorrectionStochastic OptimizationContrastive Latent Variable Models for Neural Text Generation
Deep latent variable models such as variational autoencoders and energy-based models are widely used for neural text generation. Most of them focus on matching the prior distribution with the posterior distribution of th…
Contrastive LearningText Generation