paper-with-me

Papers

Corpus for Coreference Resolution on Scientific Papers

2014-05-01 · LREC 2014 5 · Panot Chaimongkol, Akiko Aizawa, Yuka Tateisi

The ever-growing number of published scientific papers prompts the need for automatic knowledge extraction to help scientists keep up with the state-of-the-art in their respective fields. To construct a good knowledge extraction system, annotated corpora in the scientific domain are required to train machine learning models. As described in this paper, we have constructed an annotated corpus for coreference resolution in multiple scientific domains, based on an existing corpus. We have modified the annotation scheme from Message Understanding Conference to better suit scientific texts. Then we applied that to the corpus. The annotated corpus is then compared with corpora in general domains in terms of distribution of resolution classes and performance of the Stanford Dcoref coreference resolver. Through these comparisons, we have demonstrated quantitatively that our manually annotated corpus differs from a general-domain corpus, which suggests deep differences between general-domain texts and scientific texts and which shows that different approaches can be made to tackle coreference resolution for general texts and scientific texts.

📄 PDF Abstract BibTeX

Code (1)

melsk125/SciCorefCorpus 공식 구현

Tasks

coreference-resolutionCoreference ResolutionOptical Character Recognition (OCR)

Similar Papers 제목 키워드 기반

Coreference Resolution in Research Papers from Multiple Domains

2021-01-04 · Arthur Brack, Daniel Uwe Müller, Anett Hoppe, Ralph Ewerth

Coreference resolution is essential for automatic text understanding to facilitate high-level information retrieval tasks such as text summarisation or question answering. Previous work indicates that the performance of …

coreference-resolutionCoreference ResolutionInformation RetrievalRetrieval+1

SciCo: Hierarchical Cross-Document Coreference for Scientific Concepts

2021-04-18 · AKBC 2021 10 · Arie Cattan, Sophie Johnson, Daniel Weld, Ido Dagan 외

Determining coreference of concept mentions across multiple documents is a fundamental task in natural language understanding. Previous work on cross-document coreference resolution (CDCR) typically considers mentions of…

coreference-resolutionCoreference ResolutionCross Document Coreference ResolutionNatural Language Understanding

Marmara Turkish Coreference Corpus and Coreference Resolution Baseline

2017-06-06 · Peter Schüller, Kübra Cıngıllı, Ferit Tunçer, Barış Gün Sürmeli 외

We describe the Marmara Turkish Coreference Corpus, which is an annotation of the whole METU-Sabanci Turkish Treebank with mentions and coreference chains. Collecting eight or more independent annotations for each docume…

coreference-resolutionCoreference Resolution

qxoRef 1.0: A coreference corpus and mention-pair baseline for coreference resolution in Conchucos Quechua

2021-06-01 · NAACL (AmericasNLP) 2021 6 · Elizabeth Pankratz

This paper introduces qxoRef 1.0, the first coreference corpus to be developed for a Quechuan language, and describes a baseline mention-pair coreference resolution system developed for this corpus. The evaluation of thi…

coreference-resolutionCoreference Resolution

SciCorp: A Corpus of English Scientific Articles Annotated for Information Status Analysis

2016-05-01 · LREC 2016 5 · Ina Roesiger

This paper presents SciCorp, a corpus of full-text English scientific papers of two disciplines, genetics and computational linguistics. The corpus comprises co-reference and bridging information as well as information s…

Articles