paper-with-me

홈 › Papers

Robust Handling of Polysemy via Sparse Representations

2018-05-18 · SEMEVAL 2018 6 · Abhijit Mahabal, Dan Roth, Sid Mittal

Words are polysemous and multi-faceted, with many shades of meanings. We suggest that sparse distributed representations are more suitable than other, commonly used, (dense) representations to express these multiple facets, and present Category Builder, a working system that, as we show, makes use of sparse representations to support multi-faceted lexical representations. We argue that the set expansion task is well suited to study these meaning distinctions since a word may belong to multiple sets with a different reason for membership in each. We therefore exhibit the performance of Category Builder on this task, while showing that our representation captures at the same time analogy problems such as "the Ganga of Egypt" or "the Voldemort of Tolkien". Category Builder is shown to be a more expressive lexical representation and to outperform dense representations such as Word2Vec in some analogy classes despite being shown only two of the three input terms.

📄 PDF Abstract BibTeX arXiv:1805.07398

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Ultra-High Dimensional Sparse Representations with Binarization for Efficient Text Retrieval

2021-04-15 · EMNLP 2021 11 · Kyoung-Rok Jang, Junmo Kang, Giwon Hong, Sung-Hyon Myaeng 외

The semantic matching capabilities of neural information retrieval can ameliorate synonymy and polysemy problems of symbolic approaches. However, neural models' dense representations are more suitable for re-ranking, due…

BinarizationInformation RetrievalLanguage ModellingRe-Ranking+3

Let's Play Mono-Poly: BERT Can Reveal Words' Polysemy Level and Partitionability into Senses

2021-04-29 · Aina Garí Soler, Marianna Apidianaki

Pre-trained language models (LMs) encode rich information about linguistic structure but their knowledge about lexical polysemy remains unclear. We propose a novel experimental setup for analysing this knowledge in LMs s…

Clustering tweets usingWikipedia concepts

2014-05-01 · LREC 2014 5 · Guoyu Tang, Yunqing Xia, Weizhi Wang, Raymond Lau 외

Two challenging issues are notable in tweet clustering. Firstly, the sparse data problem is serious since no tweet can be longer than 140 characters. Secondly, synonymy and polysemy are rather common because users intend…

ClusteringText Clustering

Gaussian Hierarchical Latent Dirichlet Allocation: Bringing Polysemy Back

2020-02-25 · Takahiro Yoshida, Ryohei Hisano, Takaaki Ohnishi

Topic models are widely used to discover the latent representation of a set of documents. The two canonical models are latent Dirichlet allocation, and Gaussian latent Dirichlet allocation, where the former uses multinom…

Topic Models

New Polysemy Structures in Wordnets Induced by Vertical Polysemy

2019-07-01 · GWC 2019 7 · Ahti Lohk, Heili Orav, Kadri Vare, Francis Bond 외

This paper aims to study auto-hyponymy and auto-troponymy relations (or vertical polysemy) in 11 wordnets uploaded into the new Open Multilingual Wordnet (OMW) webpage. We investigate how vertical polysemy forms polysemy…