paper-with-me

홈 › Papers

Language Resources and Annotation Tools for Cross-Sentence Relation Extraction

2014-05-01 · LREC 2014 5 · Sebastian Krause, Hong Li, Feiyu Xu, Hans Uszkoreit, Robert Hummel, Luise Spielhagen

In this paper, we present a novel combination of two types of language resources dedicated to the detection of relevant relations (RE) such as events or facts across sentence boundaries. One of the two resources is the sar-graph, which aggregates for each target relation ten thousands of linguistic patterns of semantically associated relations that signal instances of the target relation (Uszkoreit and Xu, 2013). These have been learned from the Web by intra-sentence pattern extraction (Krause et al., 2012) and after semantic filtering and enriching have been automatically combined into a single graph. The other resource is cockrACE, a specially annotated corpus for the training and evaluation of cross-sentence RE. By employing our powerful annotation tool Recon, annotators mark selected entities and relations (including events), coreference relations among these entities and events, and also terms that are semantically related to the relevant relations and events. This paper describes how the two resources are created and how they complement each other.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Coreference ResolutionDependency ParsingRelationRelation ExtractionSentence

Similar Papers 제목 키워드 기반

A Basic Language Resource Kit for Persian

2012-05-01 · LREC 2012 5 · Mojgan Seraji, Be{\'a}ta Megyesi, Joakim Nivre

Persian with its about 100,000,000 speakers in the world belongs to the group of languages with less developed linguistically annotated resources and tools. The few existing resources and tools are neither open source no…

Part-Of-Speech TaggingPOSSentenceSentence segmentation+1

Guidelines for Fine-grained Sentence-level Arabic Readability Annotation

2024-10-11 · Nizar Habash, Hanada Taha-Thomure, Khalid N. Elmadani, Zeina Zeino 외

This paper presents the foundational framework and initial findings of the Balanced Arabic Readability Evaluation Corpus (BAREC) project, designed to address the need for comprehensive Arabic language resources aligned w…

BenchmarkingSentence

Cross-lingual Multi-Level Adversarial Transfer to Enhance Low-Resource Name Tagging

2019-06-01 · NAACL 2019 6 · Lifu Huang, Heng Ji, Jonathan May

We focus on improving name tagging for low-resource languages using annotations from related languages. Previous studies either directly project annotations from a source language to a target language using cross-lingual…

Cross-Lingual TransferSentence

Cross-lingual annotation: a road map for low- and no-resource languages

2020-12-01 · DMR (COLING) 2020 12 · Meagan Vigus, Jens E. L. Van Gysel, Tim O’Gorman, Andrew Cowell 외

This paper presents a “road map” for the annotation of semantic categories in typologically diverse languages, with potentially few linguistic resources, and often no existing computational resources. Past semantic annot…

Sentence

GATEtoGerManC: A GATE-based Annotation Pipeline for Historical German

2012-05-01 · LREC 2012 5 · Silke Scheible, Richard J. Whitt, Martin Durrell, Paul Bennett

We describe a new GATE-based linguistic annotation pipeline for Early Modern German, which can be used to annotate historical texts with word tokens, sentence boundaries, lemmas, and POS tags. The pipeline is based on a …

POSPOS TaggingSentence