paper-with-me

홈 › Papers

Why So Down? The Role of Negative (and Positive) Pointwise Mutual Information in Distributional Semantics

2019-08-19 · Alexandre Salle, Aline Villavicencio

In distributional semantics, the pointwise mutual information ($\mathit{PMI}$) weighting of the cooccurrence matrix performs far better than raw counts. There is, however, an issue with unobserved pair cooccurrences as $\mathit{PMI}$ goes to negative infinity. This problem is aggravated by unreliable statistics from finite corpora which lead to a large number of such pairs. A common practice is to clip negative $\mathit{PMI}$ ($\mathit{\texttt{-} PMI}$) at $0$, also known as Positive $\mathit{PMI}$ ($\mathit{PPMI}$). In this paper, we investigate alternative ways of dealing with $\mathit{\texttt{-} PMI}$ and, more importantly, study the role that negative information plays in the performance of a low-rank, weighted factorization of different $\mathit{PMI}$ matrices. Using various semantic and syntactic tasks as probes into models which use either negative or positive $\mathit{PMI}$ (or both), we find that most of the encoded semantics and syntax come from positive $\mathit{PMI}$, in contrast to $\mathit{\texttt{-} PMI}$ which contributes almost exclusively syntactic information. Our findings deepen our understanding of distributional semantics, while also introducing novel $PMI$ variants and grounding the popular $PPMI$ measure.

📄 PDF Abstract BibTeX arXiv:1908.06941

Code (1)

alexandres/lexvec 공식 구현

Similar Papers 제목 키워드 기반

Fine-tuning Distributional Semantic Models for Closely-Related Languages

2021-04-01 · EACL (VarDial) 2021 4 · Kushagra Bhatia, Divyanshu Aggarwal, Ashwini Vaidya

In this paper we compare the performance of three models: SGNS (skip-gram negative sampling) and augmented versions of SVD (singular value decomposition) and PPMI (Positive Pointwise Mutual Information) on a word similar…

Word Similarity

Enhancing the LexVec Distributed Word Representation Model Using Positional Contexts and External Memory

2016-06-03 · Alexandre Salle, Marco Idiart, Aline Villavicencio

In this paper we take a state-of-the-art model for distributed word representation that explicitly factorizes the positive pointwise mutual information (PPMI) matrix using window sampling and negative sampling and addres…

Word Similarity

Projection Al\'eatoire Non-N\'egative pour le Calcul de Word Embedding / Non-Negative Randomized Word Embedding

2017-06-01 · JEPTALNRECITAL 2017 6 · Behrang Qasemizadeh, Laura Kallmeyer, Aurelie Herbelot

Non-Negative Randomized Word Embedding We propose a word embedding method which is based on a novel random projection technique. We show that weighting methods such as positive pointwise mutual information (PPMI) can be …

Relevant and Informative Response Generation using Pointwise Mutual Information

2019-08-01 · WS 2019 8 · Junya Takayama, Yuki Arase

A sequence-to-sequence model tends to generate generic responses with little information for input utterances. To solve this problem, we propose a neural model that generates relevant and informative responses. Our model…

DecoderResponse Generation

Pre-training-free Image Manipulation Localization through Non-Mutually Exclusive Contrastive Learning

2023-09-26 · ICCV 2023 1 · Jizhe Zhou, Xiaochen Ma, Xia Du, Ahmed Y. Alhammadi 외

Deep Image Manipulation Localization (IML) models suffer from training data insufficiency and thus heavily rely on pre-training. We argue that contrastive learning is more suitable to tackle the data insufficiency proble…

Contrastive LearningImage ManipulationImage Manipulation Localization