paper-with-me

홈 › Papers

Learning an arbitrary mixture of two multinomial logits

2020-07-01 · Wenpin Tang

In this paper, we consider mixtures of multinomial logistic models (MNL), which are known to $\epsilon$-approximate any random utility model. Despite its long history and broad use, rigorous results are only available for learning a uniform mixture of two MNLs. Continuing this line of research, we study the problem of learning an arbitrary mixture of two MNLs. We show that the identifiability of the mixture models may only fail on an algebraic variety of a negligible measure. This is done by reducing the problem of learning a mixture of two MNLs to the problem of solving a system of univariate quartic equations. We also devise an algorithm to learn any mixture of two MNLs using a polynomial number of samples and a linear number of queries, provided that a mixture of two MNLs over some finite universe is identifiable. Several numerical experiments and conjectures are also presented.

📄 PDF Abstract BibTeX arXiv:2007.00204

Code (0)

등록된 구현이 없습니다.

Tasks

Vocal Bursts Valence Prediction

Similar Papers 제목 키워드 기반

Learning a Mixture of Two Multinomial Logits

2018-07-01 · ICML 2018 7 · Flavio Chierichetti, Ravi Kumar, Andrew Tomkins

The classical Multinomial Logit (MNL) is a behavioral model for user choice. In this model, a user is offered a slate of choices (a subset of a finite universe of $n$ items), and selects exactly one item from the sl…

Vocal Bursts Valence Prediction

The statistical Minkowski distances: Closed-form formula for Gaussian Mixture Models

2019-01-09 · Frank Nielsen

The traditional Minkowski distances are induced by the corresponding Minkowski norms in real-valued vector spaces. In this work, we propose novel statistical symmetric distances based on the Minkowski's inequality for pr…

DiversityForm

Better Conditional Density Estimation for Neural Networks

2016-06-07 · Wesley Tansey, Karl Pichotta, James G. Scott

The vast majority of the neural network literature focuses on predicting point values for a given set of response variables, conditioned on a feature vector. In many cases we need to model the full joint conditional dist…

Density Estimation

Asymptotic Behavior of Bayesian Generalization Error in Multinomial Mixtures

2022-03-14 · Takumi Watanabe, Sumio Watanabe

Multinomial mixtures are widely used in the information engineering field, however, their mathematical properties are not yet clarified because they are singular learning models. In fact, the models are non-identifiable …

Learning large softmax mixtures with warm start EM

2024-09-16 · Xin Bing, Florentina Bunea, Jonathan Niles-Weed, Marten Wegkamp

Mixed multinomial logits are discrete mixtures introduced several decades ago to model the probability of choosing an attribute from $p$ possible candidates, in heterogeneous populations. The model has recently attracted…

Attributeparameter estimation