paper-with-me

Papers

Span Model for Open Information Extraction on Accurate Corpus

2019-01-30 · Junlang Zhan, Hai Zhao

Open information extraction (Open IE) is a challenging task especially due to its brittle data basis. Most of Open IE systems have to be trained on automatically built corpus and evaluated on inaccurate test set. In this work, we first alleviate this difficulty from both sides of training and test sets. For the former, we propose an improved model design to more sufficiently exploit training dataset. For the latter, we present our accurately re-annotated benchmark test set (Re-OIE6) according to a series of linguistic observation and analysis. Then, we introduce a span model instead of previous adopted sequence labeling formulization for n-ary Open IE. Our newly introduced model achieves new state-of-the-art performance on both benchmark evaluation datasets.

📄 PDF Abstract BibTeX arXiv:1901.10879

Code (1)

zhanjunlang/Span_OIE 공식 구현 pytorch

Tasks

Open Information Extraction

Similar Papers 제목 키워드 기반

PubMedCausal: A Span-Level Annotated Corpus for Causal Relation Extraction in Biomedical Text

2026-05-27 · Ifeoluwa Kunle-John, Josiah Paul, Oluwatosin Agbaakin, Peter Aina 외 arxiv

Causal relation extraction (CRE) is central to biomedical text mining, but current resources often conflate causal relations with broader associations, restrict annotation to sentence-level examples, or focus mainly on e…

Relation ExtractionTransfer Learning

The CONCISUS Corpus of Event Summaries

2012-05-01 · LREC 2012 5 · Horacio Saggion, S Szasz, ra

Text summarization and information extraction systems require adaptation to new domains and languages. This adaptation usually depends on the availability of language resources such as corpora. In this paper we present a…

Text GenerationText Summarization

MuLMS: A Multi-Layer Annotated Text Corpus for Information Extraction in the Materials Science Domain

2023-10-24 · Timo Pierre Schrader, Matteo Finco, Stefan Grünewald, Felix Hildebrand 외

Keeping track of all relevant recent publications and experimental results for a research area is a challenging task. Prior work has demonstrated the efficacy of information extraction models in various scientific areas.…

Articles

An Annotated Corpus of Emerging Anglicisms in Spanish Newspaper Headlines

2020-04-06 · Elena Álvarez-Mellado

The extraction of anglicisms (lexical borrowings from English) is relevant both for lexicographic purposes and for NLP downstream tasks. We introduce a corpus of European Spanish newspaper headlines annotated with anglic…

An Annotated Corpus of Emerging Anglicisms in Spanish Newspaper Headlines

2020-05-01 · LREC 2020 5 · Elena Alvarez-Mellado

The extraction of anglicisms (lexical borrowings from English) is relevant both for lexicographic purposes and for NLP downstream tasks. We introduce a corpus of European Spanish newspaper headlines annotated with anglic…