paper-with-me

Papers

Distributional Thesauri for Information Retrieval and vice versa

2016-05-01 · LREC 2016 5 · Vincent Claveau, Ewa Kijak

Distributional thesauri are useful in many tasks of Natural Language Processing. In this paper, we address the problem of building and evaluating such thesauri with the help of Information Retrieval (IR) concepts. Two main contributions are proposed. First, following the work of [8], we show how IR tools and concepts can be used with success to build a thesaurus. Through several experiments and by evaluating directly the results with reference lexicons, we show that some IR models outperform state-of-the-art systems. Secondly, we use IR as an applicative framework to indirectly evaluate the generated thesaurus. Here again, this task-based evaluation validates the IR approach used to build the thesaurus. Moreover, it allows us to compare these results with those from the direct evaluation framework used in the literature. The observed differences bring these evaluation habits into question.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Information RetrievalRetrieval

Similar Papers 제목 키워드 기반

Direct vs. indirect evaluation of distributional thesauri

2016-12-01 · COLING 2016 12 · Vincent Claveau, Ewa Kijak

With the success of word embedding methods in various Natural Language Processing tasks, all the field of distributional semantics has experienced a renewed interest. Beside the famous word2vec, recent studies have prese…

Information RetrievalRetrieval

Word Embedding-based Antonym Detection using Thesauri and Distributional Information

2015-05-01 · HLT 2015 5 · Yutaka Sasaki, Makoto Miwa, Masataka Ono
Dependency ParsingLearning Word Embeddingsnamed-entity-recognitionNamed Entity Recognition+5

Comparing Similarity Measures for Distributional Thesauri

2014-05-01 · LREC 2014 5 · Muntsa Padr{\'o}, Marco Idiart, Aline Villavicencio, Carlos Ramisch

Distributional thesauri have been applied for a variety of tasks involving semantic relatedness. In this paper, we investigate the impact of three parameters: similarity measures, frequency thresholds and association sco…

Dimensionality Reduction

Distributed Distributional Similarities of Google Books Over the Centuries

2014-05-01 · LREC 2014 5 · Martin Riedl, Richard Steuer, Chris Biemann

This paper introduces a distributional thesaurus and sense clusters computed on the complete Google Syntactic N-grams, which is extracted from Google Books, a very large corpus of digitized books published between 1520 a…

Graph Clustering

Extrinsic Evaluation of French Dependency Parsers on a Specialized Corpus: Comparison of Distributional Thesauri

2020-05-01 · LREC 2020 5 · Ludovic Tanguy, Pauline Brunet, Olivier Ferret

We present a study in which we compare 11 different French dependency parsers on a specialized corpus (consisting of research articles on NLP from the proceedings of the TALN conference). Due to the lack of a suitable go…

Articles