Language Resources and Annotation Tools for Cross-Sentence Relation Extraction
In this paper, we present a novel combination of two types of language resources dedicated to the detection of relevant relations (RE) such as events or facts across sentence boundaries. One of the two resources is the sar-graph, which aggregates for each target relation ten thousands of linguistic patterns of semantically associated relations that signal instances of the target relation (Uszkoreit and Xu, 2013). These have been learned from the Web by intra-sentence pattern extraction (Krause et al., 2012) and after semantic filtering and enriching have been automatically combined into a single graph. The other resource is cockrACE, a specially annotated corpus for the training and evaluation of cross-sentence RE. By employing our powerful annotation tool Recon, annotators mark selected entities and relations (including events), coreference relations among these entities and events, and also terms that are semantically related to the relevant relations and events. This paper describes how the two resources are created and how they complement each other.
Code (0)
등록된 구현이 없습니다.
Tasks
Coreference ResolutionDependency ParsingRelationRelation ExtractionSentenceSimilar Papers 제목 키워드 기반
A Basic Language Resource Kit for Persian
Persian with its about 100,000,000 speakers in the world belongs to the group of languages with less developed linguistically annotated resources and tools. The few existing resources and tools are neither open source no…
Part-Of-Speech TaggingPOSSentenceSentence segmentation+1Guidelines for Fine-grained Sentence-level Arabic Readability Annotation
This paper presents the foundational framework and initial findings of the Balanced Arabic Readability Evaluation Corpus (BAREC) project, designed to address the need for comprehensive Arabic language resources aligned w…
BenchmarkingSentenceCross-lingual Multi-Level Adversarial Transfer to Enhance Low-Resource Name Tagging
We focus on improving name tagging for low-resource languages using annotations from related languages. Previous studies either directly project annotations from a source language to a target language using cross-lingual…
Cross-Lingual TransferSentenceCross-lingual annotation: a road map for low- and no-resource languages
This paper presents a “road map” for the annotation of semantic categories in typologically diverse languages, with potentially few linguistic resources, and often no existing computational resources. Past semantic annot…
SentenceGATEtoGerManC: A GATE-based Annotation Pipeline for Historical German
We describe a new GATE-based linguistic annotation pipeline for Early Modern German, which can be used to annotate historical texts with word tokens, sentence boundaries, lemmas, and POS tags. The pipeline is based on a …
POSPOS TaggingSentence