paper-with-me

Papers

Code Book for the Annotation of Diverse Cross-Document Coreference of Entities in News Articles

2023-10-18 · Jakob Vogel

This paper presents a scheme for annotating coreference across news articles, extending beyond traditional identity relations by also considering near-identity and bridging relations. It includes a precise description of how to set up Inception, a respective annotation tool, how to annotate entities in news articles, connect them with diverse coreferential relations, and link them across documents to Wikidata's global knowledge graph. This multi-layered annotation approach is discussed in the context of the problem of media bias. Our main contribution lies in providing a methodology for creating a diverse cross-document coreference corpus which can be applied to the analysis of media bias by word-choice and labelling.

📄 PDF Abstract BibTeX arXiv:2310.12064

Code (0)

등록된 구현이 없습니다.

Tasks

Articles

Methods 이 논문이 사용한 방법론

SET Dynamic Sparse Training method where weight mask is updated randomly periodically

Similar Papers 제목 키워드 기반

OmniDocBench: Benchmarking Diverse PDF Document Parsing with Comprehensive Annotations

2024-12-10 · CVPR 2025 1 · Linke Ouyang, Yuan Qu, Hongbin Zhou, Jiawei Zhu 외

Document content extraction is a critical task in computer vision, underpinning the data needs of large language models (LLMs) and retrieval-augmented generation (RAG) systems. Despite recent progress, current document p…

AttributeBenchmarkingDiversityRAG+1

Diverse Word Choices, Same Reference: Annotating Lexically-Rich Cross-Document Coreference

2026-02-19 · Anastasia Zhukova, Felix Hamborg, Karsten Donnay, Norman Meuschke 외 arxiv

Cross-document coreference resolution (CDCR) identifies and links mentions of the same entities and events across related documents, enabling content analysis that aggregates information at the level of discourse partici…

Coreference Resolution

M6Doc: A Large-Scale Multi-Format, Multi-Type, Multi-Layout, Multi-Language, Multi-Annotation Category Dataset for Modern Document Layout Analysis

2023-01-01 · CVPR 2023 1 · Hiuyi Cheng, Peirong Zhang, Sihang Wu, Jiaxin Zhang 외

Document layout analysis is a crucial prerequisite for document understanding, including document retrieval and conversion. Most public datasets currently contain only PDF documents and lack realistic documents. Mode…

ArticlesDocument Layout Analysisdocument understandingInstance Segmentation+2

M$^{6}$Doc: A Large-Scale Multi-Format, Multi-Type, Multi-Layout, Multi-Language, Multi-Annotation Category Dataset for Modern Document Layout Analysis

2023-05-15 · Hiuyi Cheng, Peirong Zhang, Sihang Wu, Jiaxin Zhang 외

Document layout analysis is a crucial prerequisite for document understanding, including document retrieval and conversion. Most public datasets currently contain only PDF documents and lack realistic documents. Models t…

ArticlesDocument Layout Analysisdocument understandingInstance Segmentation+2

BOOKCOREF: Coreference Resolution at Book Scale

2025-07-16 · Giuliano Martinelli, Tommaso Bonomo, Pere-Lluís Huguet Cabot, Roberto Navigli arxiv

Coreference Resolution systems are typically evaluated on benchmarks containing small- to medium-scale documents. When it comes to evaluating long texts, however, existing benchmarks, such as LitBank, remain limited in l…

Coreference Resolution