paper-with-me

Papers

Factorial LDA: Sparse Multi-Dimensional Text Models

2012-12-01 · NeurIPS 2012 12 · Michael Paul, Mark Dredze

Multi-dimensional latent variable models can capture the many latent factors in a text corpus, such as topic, author perspective and sentiment. We introduce factorial LDA, a multi-dimensional latent variable model in which a document is influenced by K different factors, and each word token depends on a K-dimensional vector of latent variables. Our model incorporates structured word priors and learns a sparse product of factors. Experiments on research abstracts show that our model can learn latent factors such as research topic, scientific discipline, and focus (e.g. methods vs. applications.) Our modeling improvements reduce test perplexity and improve human interpretability of the discovered factors.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

LDA Linear discriminant analysis (LDA), normal discriminant analysis (NDA), or discriminant function analysis is a generalization of Fisher's linear discriminant, a method used in…

Similar Papers 제목 키워드 기반

Sparse Estimation Using General Likelihoods and Non-Factorial Priors

2009-12-01 · NeurIPS 2009 12 · David P. Wipf, Srikantan S. Nagarajan

Finding maximally sparse representations from overcomplete feature dictionaries frequently involves minimizing a cost function composed of a likelihood (or data fit) term and a prior (or penalty function) that favors spa…

feature selectionGeneral Classification

Decoupled Learning for Factorial Marked Temporal Point Processes

2018-01-21 · Weichang Wu, Junchi Yan, Xiaokang Yang, Hongyuan Zha

This paper introduces the factorial marked temporal point process model and presents efficient learning methods. In conventional (multi-dimensional) marked temporal point process models, event is often encoded by a singl…

Point Processes

Learning with Hidden Factorial Structure

2024-11-02 · Charles Arnal, Clement Berenfeld, Simon Rosenberg, Vivien Cabannes

Statistical learning in high-dimensional spaces is challenging without a strong underlying data structure. Recent advances with foundational models suggest that text and image data contain such hidden structures, which h…

Exploiting locality in high-dimensional factorial hidden Markov models

2019-02-05 · Lorenzo Rimella, Nick Whiteley

We propose algorithms for approximate filtering and smoothing in high-dimensional Factorial hidden Markov models. The approximation involves discarding, in a principled way, likelihood factors according to a notion of lo…

Vocal Bursts Intensity Prediction

Sparse, complex-valued representations of natural sounds learned with phase and amplitude continuity priors

2013-12-17 · Wiktor Mlynarski

Complex-valued sparse coding is a data representation which employs a dictionary of two-dimensional subspaces, while imposing a sparse, factorial prior on complex amplitudes. When trained on a dataset of natural image pa…