paper-with-me

홈 › Papers

Keywords lie far from the mean of all words in local vector space

2020-08-21 · Eirini Papagiannopoulou, Grigorios Tsoumakas, Apostolos N. Papadopoulos

Keyword extraction is an important document process that aims at finding a small set of terms that concisely describe a document's topics. The most popular state-of-the-art unsupervised approaches belong to the family of the graph-based methods that build a graph-of-words and use various centrality measures to score the nodes (candidate keywords). In this work, we follow a different path to detect the keywords from a text document by modeling the main distribution of the document's words using local word vector representations. Then, we rank the candidates based on their position in the text and the distance between the corresponding local vectors and the main distribution's center. We confirm the high performance of our approach compared to strong baselines and state-of-the-art unsupervised keyword extraction methods, through an extended experimental study, investigating the properties of the local representations.

📄 PDF Abstract BibTeX arXiv:2008.09513

Code (1)

epapagia/LocalVectors_AKE 공식 구현

Tasks

AllKeyword ExtractionPosition

Similar Papers 제목 키워드 기반

A Generalized Vector Space Model for Ontology-Based Information Retrieval

2018-07-20 · Vuong M. Ngo, Tru H. Cao

Named entities (NE) are objects that are referred to by names such as people, organizations and locations. Named entities and keywords are important to the meaning of a document. We propose a generalized vector space mod…

Information RetrievalRetrieval

Exploring Combinations of Ontological Features and Keywords for Text Retrieval

2018-07-20 · Cao Tru H., Le Khanh C., Ngo Vuong M.

Named entities have been considered and combined with keywords to enhance information retrieval performance. However, there is not yet a formal and complete model that takes into account entity names, classes, and identi…

Information RetrievalRetrievalText Retrieval

Information Retrieval in long documents: Word clustering approach for improving Semantics

2023-02-20 · Paul Mbate Mekontchou, Armel Fotsoh, Bernabe Batchakui, Eddy Ella

In this paper, we propose an alternative to deep neural networks for semantic information retrieval for the case of long documents. This new approach exploiting clustering techniques to take into account the meaning of w…

ClusteringInformation RetrievalRetrieval

Searching for Discriminative Words in Multidimensional Continuous Feature Space

2022-11-26 · Marius Sajgalik, Michal Barla, Maria Bielikova

Word feature vectors have been proven to improve many NLP tasks. With recent advances in unsupervised learning of these feature vectors, it became possible to train it with much more data, which also resulted in better q…

Part-Of-Speech Tagging

RPM-Oriented Query Rewriting Framework for E-commerce Keyword-Based Sponsored Search

2019-10-28 · Xiuying Chen, Daorui Xiao, Shen Gao, Guojun Liu 외

Sponsored search optimizes revenue and relevance, which is estimated by Revenue Per Mille (RPM). Existing sponsored search models are all based on traditional statistical models, which have poor RPM performance when quer…