paper-with-me

홈 › Papers

Extended Named Entities Annotation on OCRed Documents: From Corpus Constitution to Evaluation Campaign

2012-05-01 · LREC 2012 5 · Olivier Galibert, Sophie Rosset, Cyril Grouin, Pierre Zweigenbaum, Ludovic Quintard

Within the framework of the Quaero project, we proposed a new definition of named entities, based upon an extension of the coverage of named entities as well as the structure of those named entities. In this new definition, the extended named entities we proposed are both hierarchical and compositional. In this paper, we focused on the annotation of a corpus composed of press archives, OCRed from French newspapers of December 1890. We present the methodology we used to produce the corpus and the characteristics of the corpus in terms of named entities annotation. This annotated corpus has been used in an evaluation campaign. We present this evaluation, the metrics we used and the results obtained by the participants.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Named Entity Recognition (NER)Optical Character Recognition (OCR)

Similar Papers 제목 키워드 기반

On the Robustness of Document-Level Relation Extraction Models to Entity Name Variations

2024-06-11 · Shiao Meng, Xuming Hu, Aiwei Liu, Fukun Ma 외

Driven by the demand for cross-sentence and large-scale relation extraction, document-level relation extraction (DocRE) has attracted increasing research interest. Despite the continuous improvement in performance, we fi…

Document-level Relation ExtractionIn-Context LearningRelationRelation Extraction+1

Old Content and Modern Tools - Searching Named Entities in a Finnish OCRed Historical Newspaper Collection 1771-1910

2016-11-09 · Kimmo Kettunen, Eetu Mäkelä, Teemu Ruokolainen, Juha Kuokkala 외

Named Entity Recognition (NER), search, classification and tagging of names and name like frequent informational elements in texts, has become a standard information extraction procedure for textual data. NER has been ap…

named-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)NER+1

Revisiting DocRED -- Addressing the False Negative Problem in Relation Extraction

2022-05-25 · Qingyu Tan, Lu Xu, Lidong Bing, Hwee Tou Ng 외

The DocRED dataset is one of the most popular and widely used benchmarks for document-level relation extraction (RE). It adopts a recommend-revise annotation scheme so as to have a large-scale annotated dataset. However,…

Document-level Relation ExtractionRelationRelation Extraction

DocRED: A Large-Scale Document-Level Relation Extraction Dataset

2019-06-14 · ACL 2019 7 · Yuan Yao, Deming Ye, Peng Li, Xu Han 외

Multiple entities in a document generally exhibit complex inter-sentence relations, and cannot be well handled by existing relation extraction (RE) methods that typically focus on extracting intra-sentence relations for …

Document-level Relation ExtractionRelationRelation ExtractionSentence

Does Recommend-Revise Produce Reliable Annotations? An Analysis on Missing Instances in DocRED

2022-04-17 · ACL 2022 5 · Quzhe Huang, Shibo Hao, Yuan Ye, Shengqi Zhu 외

DocRED is a widely used dataset for document-level relation extraction. In the large-scale annotation, a \textit{recommend-revise} scheme is adopted to reduce the workload. Within this scheme, annotators are provided wit…

Document-level Relation ExtractionRelationRelation Extraction