paper-with-me

홈 › Papers

Stochastic Variational Inference with Tuneable Stochastic Annealing

2025-04-04 · John Paisley, Ghazal Fazelnia, Brian Barr

In this paper, we exploit the observation that stochastic variational inference (SVI) is a form of annealing and present a modified SVI approach -- applicable to both large and small datasets -- that allows the amount of annealing done by SVI to be tuned. We are motivated by the fact that, in SVI, the larger the batch size the more approximately Gaussian is the intrinsic noise, but the smaller its variance. This low variance reduces the amount of annealing which is needed to escape bad local optimal solutions. We propose a simple method for achieving both goals of having larger variance noise to escape bad local optimal solutions and more data information to obtain more accurate gradient directions. The idea is to set an actual batch size, which may be the size of the data set, and a smaller effective batch size that matches the larger level of variance at this smaller batch size. The result is an approximation to the maximum entropy stochastic gradient at this variance level. We theoretically motivate our approach for the framework of conjugate exponential family models and illustrate the method empirically on the probabilistic matrix factorization collaborative filter, the Latent Dirichlet Allocation topic model, and the Gaussian mixture model.

📄 PDF Abstract BibTeX arXiv:2504.03902

Code (0)

등록된 구현이 없습니다.

Tasks

Variational Inference

Methods 이 논문이 사용한 방법론

Variational Inference 설명 없음
SET Dynamic Sparse Training method where weight mask is updated randomly periodically

Similar Papers 제목 키워드 기반

Stochastic Annealing for Variational Inference

2015-05-25 · San Gultekin, Aonan Zhang, John Paisley

We empirically evaluate a stochastic annealing strategy for Bayesian posterior optimization with variational inference. Variational inference is a deterministic approach to approximate posterior inference in Bayesian mod…

Variational Inference

Balancing Two-Player Stochastic Games with Soft Q-Learning

2018-02-09 · Jordi Grau-Moya, Felix Leibfried, Haitham Bou-Ammar

Within the context of video games the notion of perfectly rational agents can be undesirable as it leads to uninteresting situations, where humans face tough adversarial decision makers. Current frameworks for stochastic…

Q-LearningReinforcement LearningReinforcement Learning (RL)Vocal Bursts Valence Prediction

Time Series Clustering with General State Space Models via Stochastic Variational Inference

2024-06-29 · Ryoichi Ishizuka, Takashi Imai, Kaoru Kawamoto

In this paper, we propose a novel method of model-based time series clustering with mixtures of general state space models (MSSMs). Each component of MSSMs is associated with each cluster. An advantage of the proposed me…

Clusteringparameter estimationState Space ModelsTime Series+2

SQ-VAE: Variational Bayes on Discrete Representation with Self-annealed Stochastic Quantization

2022-05-16 · Yuhta Takida, Takashi Shibuya, WeiHsiang Liao, Chieh-Hsin Lai 외

One noted issue of vector-quantized variational autoencoder (VQ-VAE) is that the learned discrete representation uses only a fraction of the full capacity of the codebook, also known as codebook collapse. We hypothesize …

Quantization

A Diffusion Approximation Theory of Momentum SGD in Nonconvex Optimization

2018-02-14 · Tianyi Liu, Zhehui Chen, Enlu Zhou, Tuo Zhao

Momentum Stochastic Gradient Descent (MSGD) algorithm has been widely applied to many nonconvex optimization problems in machine learning, e.g., training deep neural networks, variational Bayesian inference, and etc. Des…

Bayesian InferenceDimensionality ReductionStochastic Optimization