paper-with-me

Papers

The Knowref Coreference Corpus: Removing Gender and Number Cues for Difficult Pronominal Anaphora Resolution

2018-11-02 · ACL 2019 7 · Ali Emami, Paul Trichelair, Adam Trischler, Kaheer Suleman, Hannes Schulz, Jackie Chi Kit Cheung

We introduce a new benchmark for coreference resolution and NLI, Knowref, that targets common-sense understanding and world knowledge. Previous coreference resolution tasks can largely be solved by exploiting the number and gender of the antecedents, or have been handcrafted and do not reflect the diversity of naturally occurring text. We present a corpus of over 8,000 annotated text passages with ambiguous pronominal anaphora. These instances are both challenging and realistic. We show that various coreference systems, whether rule-based, feature-rich, or neural, perform significantly worse on the task than humans, who display high inter-annotator agreement. To explain this performance gap, we show empirically that state-of-the art models often fail to capture context, instead relying on the gender or number of candidate antecedents to make a decision. We then use problem-specific insights to propose a data-augmentation trick called antecedent switching to alleviate this tendency in models. Finally, we show that antecedent switching yields promising results on other tasks as well: we use it to achieve state-of-the-art results on the GAP coreference task.

📄 PDF Abstract BibTeX arXiv:1811.01747

Code (1)

aemami1/KnowRef 공식 구현

Tasks

Common Sense Reasoningcoreference-resolutionCoreference ResolutionData AugmentationDiversityWorld Knowledge

Similar Papers 제목 키워드 기반

Polish Coreference Corpus in Numbers

2014-05-01 · LREC 2014 5 · Maciej Ogrodniczuk, Mateusz Kope{\'c}, Agata Savary

This paper attempts a preliminary interpretation of the occurrence of different types of linguistic constructs in the manually-annotated Polish Coreference Corpus by providing analyses of various statistical properties r…

Clusteringcoreference-resolutionCoreference ResolutionLemmatization+1

ANCOR\_Centre, a large free spoken French coreference corpus: description of the resource and reliability measures

2014-05-01 · LREC 2014 5 · Judith Muzerelle, Ana{\"\i}s Lefeuvre, Emmanuel Schang, Jean-Yves Antoine 외

This article presents ANCOR{\_}Centre, a French coreference corpus, available under the Creative Commons Licence. With a size of around 500,000 words, the corpus is large enough to serve the needs of data-driven approach…

Coreference ResolutionEntity LinkingInformation Retrieval

Gender Bias in Coreference Resolution: Evaluation and Debiasing Methods

2018-04-18 · NAACL 2018 6 · Jieyu Zhao, Tianlu Wang, Mark Yatskar, Vicente Ordonez 외

We introduce a new benchmark, WinoBias, for coreference resolution focused on gender bias. Our corpus contains Winograd-schema style sentences with entities corresponding to people referred by their occupation (e.g. the …

coreference-resolutionCoreference ResolutionData Augmentation

An Analysis of Dataset Overlap on Winograd-Style Tasks

2020-11-09 · COLING 2020 8 · Ali Emami, Adam Trischler, Kaheer Suleman, Jackie Chi Kit Cheung

The Winograd Schema Challenge (WSC) and variants inspired by it have become important benchmarks for common-sense reasoning (CSR). Model performance on the WSC has quickly progressed from chance-level to near-human using…

Common Sense Reasoning

Collecting a Large-Scale Gender Bias Dataset for Coreference Resolution and Machine Translation

2021-09-08 · Findings (EMNLP) 2021 11 · Shahar Levy, Koren Lazar, Gabriel Stanovsky

Recent works have found evidence of gender bias in models of machine translation and coreference resolution using mostly synthetic diagnostic datasets. While these quantify bias in a controlled experiment, they often do …

coreference-resolutionCoreference ResolutionDiagnosticMachine Translation+1