paper-with-me

Papers

BreGMN: scaled-Bregman Generative Modeling Networks

2019-06-01 · Akash Srivastava, Kristjan Greenewald, Farzaneh Mirzazadeh

The family of f-divergences is ubiquitously applied to generative modeling in order to adapt the distribution of the model to that of the data. Well-definedness of f-divergences, however, requires the distributions of the data and model to overlap completely in every time step of training. As a result, as soon as the support of distributions of data and model contain non-overlapping portions, gradient based training of the corresponding model becomes hopeless. Recent advances in generative modeling are full of remedies for handling this support mismatch problem: key ideas include either modifying the objective function to integral probability measures (IPMs) that are well-behaved even on disjoint probabilities, or optimizing a well-behaved variational lower bound instead of the true objective. We, on the other hand, establish that a complete change of the objective function is unnecessary, and instead an augmentation of the base measure of the problematic divergence can resolve the issue. Based on this observation, we propose a generative model which leverages the class of Scaled Bregman Divergences and generalizes both f-divergences and Bregman divergences. We analyze this class of divergences and show that with the appropriate choice of base measure it can resolve the support mismatch problem and incorporate geometric information. Finally, we study the performance of the proposed method and demonstrate promising results on MNIST, CelebA and CIFAR-10 datasets.

📄 PDF Abstract BibTeX arXiv:1906.00313

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

A scaled Bregman theorem with applications

2016-07-01 · NeurIPS 2016 12 · Richard Nock, Aditya Krishna Menon, Cheng Soon Ong

Bregman divergences play a central role in the design and analysis of a range of machine learning algorithms. This paper explores the use of Bregman divergences to establish reductions between such algorithms and their a…

BIG-bench Machine LearningClustering

Beta Diffusion

2023-09-14 · NeurIPS 2023 11 · Mingyuan Zhou, Tianqi Chen, Zhendong Wang, Huangjie Zheng

We introduce beta diffusion, a novel generative modeling method that integrates demasking and denoising to generate data within bounded ranges. Using scaled and shifted beta distributions, beta diffusion utilizes multipl…

Denoising

A note on the quasiconvex Jensen divergences and the quasiconvex Bregman divergences derived thereof

2019-09-19 · Frank Nielsen, Gaëtan Hadjeres

We first introduce the class of strictly quasiconvex and strictly quasiconcave Jensen divergences which are oriented (asymmetric) distances, and study some of their properties. We then define the strictly quasiconvex Bre…

Calibeating for general proper losses: A Bregman divergence approach

2026-05-17 · Maximilian Fichtl, Cristóbal Guzmán, Nishant A. Mehta arxiv

This work introduces a general framework for calibeating based on regret minimization. As compared to Foster and Hart's seminal calibeating work which had specialized treatments of Brier score (squared loss) and log loss…

Representation Learning of Compositional Data

2018-12-01 · NeurIPS 2018 12 · Marta Avalos, Richard Nock, Cheng Soon Ong, Julien Rouar 외

We consider the problem of learning a low dimensional representation for compositional data. Compositional data consists of a collection of nonnegative data that sum to a constant value. Since the parts of the collection…

Representation Learning