paper-with-me

Papers

Linked Open Data and Web Corpus Data for noun compound bracketing

2014-05-01 · LREC 2014 5 · Pierre Andr{\'e} M{\'e}nard, Caroline Barri{\`e}re

This research provides a comparison of a linked open data resource (DBpedia) and web corpus data resources (Google Web Ngrams and Google Books Ngrams) for noun compound bracketing. Large corpus statistical analysis has often been used for noun compound bracketing, and our goal is to introduce a linked open data (LOD) resource for such task. We show its particularities and its performance on the task. Results obtained on resources tested individually are promising, showing a potential for DBpedia to be included in future hybrid systems.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

SemSim: Resources for Normalized Semantic Similarity Computation Using Lexical Networks

2012-05-01 · LREC 2012 5 · Elias Iosif, Alex Potamianos, ros

We investigate the creation of corpora from web-harvested data following a scalable approach that has linear query complexity. Individual web queries are posed for a lexicon that includes thousands of nouns and the retri…

Semantic SimilaritySemantic Textual SimilarityText Categorization

Boosting Open Information Extraction with Noun-Based Relations

2014-05-01 · LREC 2014 5 · Clarissa Xavier, Vera Lima

Open Information Extraction (Open IE) is a strategy for learning relations from texts, regardless the domain and without predefining these relations. Work in this area has focused mainly on verbal relations. In order to …

Open Information ExtractionRelation Extraction

Porting Elements of the Austrian Baroque Corpus onto the Linguistic Linked Open Data Format

2013-09-01 · WS 2013 9 · Ulrike Czeitschner, Thierry Declerck, Claudia Resch

Compiling Czech Parliamentary Stenographic Protocols into a Corpus

2020-05-01 · LREC 2020 5 · Barbora Hladka, Maty{\'a}{\v{s}} Kopp, Pavel Stra{\v{n}}{\'a}k

The Parliament of the Czech Republic consists of two chambers: the Chamber of Deputies (Lower House) and the Senate (Upper House). In our work, we focus on agenda and documents that relate to the Chamber of Deputies excl…

GeoVectors: A Linked Open Corpus of OpenStreetMap Embeddings on World Scale

2021-08-30 · Nicolas Tempelmeier, Simon Gottschalk, Elena Demidova

OpenStreetMap (OSM) is currently the richest publicly available information source on geographic entities (e.g., buildings and roads) worldwide. However, using OSM entities in machine learning models and other applicatio…

BIG-bench Machine LearningEntity EmbeddingsKnowledge Graphs