paper-with-me

Papers

AnnIE: An Annotation Platform for Constructing Complete Open Information Extraction Benchmark

2021-09-15 · ACL 2022 5 · Niklas Friedrich, Kiril Gashteovski, Mingying Yu, Bhushan Kotnis, Carolin Lawrence, Mathias Niepert, Goran Glavaš

Open Information Extraction (OIE) is the task of extracting facts from sentences in the form of relations and their corresponding arguments in schema-free manner. Intrinsic performance of OIE systems is difficult to measure due to the incompleteness of existing OIE benchmarks: the ground truth extractions do not group all acceptable surface realizations of the same fact that can be extracted from a sentence. To measure performance of OIE systems more realistically, it is necessary to manually annotate complete facts (i.e., clusters of all acceptable surface realizations of the same fact) from input sentences. We propose AnnIE: an interactive annotation platform that facilitates such challenging annotation tasks and supports creation of complete fact-oriented OIE evaluation benchmarks. AnnIE is modular and flexible in order to support different use case scenarios (i.e., benchmarks covering different types of facts). We use AnnIE to build two complete OIE benchmarks: one with verb-mediated facts and another with facts encompassing named entities. Finally, we evaluate several OIE systems on our complete benchmarks created with AnnIE. Our results suggest that existing incomplete benchmarks are overly lenient, and that OIE systems are not as robust as previously reported. We publicly release AnnIE under non-restrictive license.

📄 PDF Abstract BibTeX arXiv:2109.07464

Code (1)

nfriedri/annie-annotation-platform 공식 구현

Tasks

Open Information ExtractionSentence

Similar Papers 제목 키워드 기반

VDCook:DIY video data cook your MLLMs

2026-03-04 · Chengwei Wu arxiv

We introduce VDCook: a self-evolving video data operating system, a configurable video data construction platform for researchers and vertical domain teams. Users initiate data requests via natural language queries and a…

Natural Language QueriesScene SegmentationVideo Retrieval

A Tight Lower Bound for the Approximation Guarantee of Higher-Order Singular Value Decomposition

2025-08-08 · Matthew Fahrbach, Mehrdad Ghadiri arxiv

We prove that the classic approximation guarantee for the higher-order singular value decomposition (HOSVD) is tight by constructing a tensor for which HOSVD achieves an approximation ratio of $N/(1+\varepsilon)$, for an…

ANNIE: Be Careful of Your Robots

2025-09-03 · Yiyang Huang, Zixuan Wang, Zishen Wan, Yapeng Tian 외 arxiv

The integration of vision-language-action (VLA) models into embodied AI (EAI) robots is rapidly advancing their ability to perform complex, long-horizon tasks in humancentric environments. However, EAI systems introduce …

Reconstructing Sepsis Trajectories from Clinical Case Reports using LLMs: the Textual Time Series Corpus for Sepsis

2025-04-12 · Shahriar Noroozizadeh, Jeremy C. Weiss

Clinical case reports and discharge summaries may be the most complete and accurate summarization of patient encounters, yet they are finalized, i.e., timestamped after the encounter. Complementary data structured stream…

Time Series

GATEtoGerManC: A GATE-based Annotation Pipeline for Historical German

2012-05-01 · LREC 2012 5 · Silke Scheible, Richard J. Whitt, Martin Durrell, Paul Bennett

We describe a new GATE-based linguistic annotation pipeline for Early Modern German, which can be used to annotate historical texts with word tokens, sentence boundaries, lemmas, and POS tags. The pipeline is based on a …

POSPOS TaggingSentence