paper-with-me

홈 › Papers

On Variational Bounds of Mutual Information

2019-05-16 · Ben Poole, Sherjil Ozair, Aaron van den Oord, Alexander A. Alemi, George Tucker

Estimating and optimizing Mutual Information (MI) is core to many problems in machine learning; however, bounding MI in high dimensions is challenging. To establish tractable and scalable objectives, recent work has turned to variational bounds parameterized by neural networks, but the relationships and tradeoffs between these bounds remains unclear. In this work, we unify these recent developments in a single framework. We find that the existing variational lower bounds degrade when the MI is large, exhibiting either high bias or high variance. To address this problem, we introduce a continuum of lower bounds that encompasses previous bounds and flexibly trades off bias and variance. On high-dimensional, controlled problems, we empirically characterize the bias and variance of the bounds and their gradients and demonstrate the effectiveness of our new bounds for estimation and representation learning.

📄 PDF Abstract BibTeX arXiv:1905.06922

Code (3)

Linear95/CLUB tf
RSMI-NE/RSMI-NE tf
karlstratos/doe pytorch

Tasks

Representation Learning

Similar Papers 제목 키워드 기반

Variational Information Maximization for Feature Selection

2016-06-09 · NeurIPS 2016 12 · Shuyang Gao, Greg Ver Steeg, Aram Galstyan

Feature selection is one of the most fundamental problems in machine learning. An extensive body of work on information-theoretic feature selection exists which is based on maximizing mutual information between subsets o…

feature selection

Bounds on mutual information of mixture data for classification tasks

2021-01-27 · Yijun Ding, Amit Ashok

The data for many classification problems, such as pattern and speech recognition, follow mixture distributions. To quantify the optimum performance for classification tasks, the Shannon mutual information is a natural i…

ClassificationGeneral Classificationspeech-recognitionSpeech Recognition

The Role of Mutual Information in Variational Classifiers

2020-10-22 · Matias Vera, Leonardo Rey Vega, Pablo Piantanida

Overfitting data is a well-known phenomenon related with the generation of a model that mimics too closely (or exactly) a particular instance of data, and may therefore fail to predict future observations reliably. In pr…

Variational Inference

Towards Consistency and Complementarity: A Multiview Graph Information Bottleneck Approach

2022-10-11 · Xiaolong Fan, Maoguo Gong, Yue Wu, Mingyang Zhang 외

The empirical studies of Graph Neural Networks (GNNs) broadly take the original node feature and adjacency relationship as singleview input, ignoring the rich information of multiple graph views. To circumvent this issue…

DEMI: Discriminative Estimator of Mutual Information

2020-10-05 · Ruizhi Liao, Daniel Moyer, Polina Golland, William M. Wells

Estimating mutual information between continuous random variables is often intractable and extremely challenging for high-dimensional data. Recent progress has leveraged neural networks to optimize variational lower boun…

Representation Learning