Linked Open Data and Web Corpus Data for noun compound bracketing
This research provides a comparison of a linked open data resource (DBpedia) and web corpus data resources (Google Web Ngrams and Google Books Ngrams) for noun compound bracketing. Large corpus statistical analysis has often been used for noun compound bracketing, and our goal is to introduce a linked open data (LOD) resource for such task. We show its particularities and its performance on the task. Results obtained on resources tested individually are promising, showing a potential for DBpedia to be included in future hybrid systems.
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
SemSim: Resources for Normalized Semantic Similarity Computation Using Lexical Networks
We investigate the creation of corpora from web-harvested data following a scalable approach that has linear query complexity. Individual web queries are posed for a lexicon that includes thousands of nouns and the retri…
Semantic SimilaritySemantic Textual SimilarityText CategorizationBoosting Open Information Extraction with Noun-Based Relations
Open Information Extraction (Open IE) is a strategy for learning relations from texts, regardless the domain and without predefining these relations. Work in this area has focused mainly on verbal relations. In order to …
Open Information ExtractionRelation ExtractionPorting Elements of the Austrian Baroque Corpus onto the Linguistic Linked Open Data Format
Compiling Czech Parliamentary Stenographic Protocols into a Corpus
The Parliament of the Czech Republic consists of two chambers: the Chamber of Deputies (Lower House) and the Senate (Upper House). In our work, we focus on agenda and documents that relate to the Chamber of Deputies excl…
GeoVectors: A Linked Open Corpus of OpenStreetMap Embeddings on World Scale
OpenStreetMap (OSM) is currently the richest publicly available information source on geographic entities (e.g., buildings and roads) worldwide. However, using OSM entities in machine learning models and other applicatio…
BIG-bench Machine LearningEntity EmbeddingsKnowledge Graphs