Cross-lingual Incongruences in the Annotation of Coreference
In the present paper, we deal with incongruences in English-German multilingual coreference annotation and present automated methods to discover them. More specifically, we automatically detect full coreference chains in parallel texts and analyse discrepancies in their annotations. In doing so, we wish to find out whether the discrepancies rather derive from language typological constraints, from the translation or the actual annotation process. The results of our study contribute to the referential analysis of similarities and differences across languages and support evaluation of cross-lingual coreference annotation. They are also useful for cross-lingual coreference resolution systems and contrastive linguistic studies.
Code (0)
등록된 구현이 없습니다.
Tasks
coreference-resolutionCoreference ResolutionTranslationSimilar Papers 제목 키워드 기반
PAWS: A Multi-lingual Parallel Treebank with Anaphoric Relations
We present PAWS, a multi-lingual parallel treebank with coreference annotation. It consists of English texts from the Wall Street Journal translated into Czech, Russian and Polish. In addition, the texts are syntacticall…
Coreference ResolutionMachine TranslationInvestigating Multilingual Coreference Resolution by Universal Annotations
Multilingual coreference resolution (MCR) has been a long-standing and challenging task. With the newly proposed multilingual coreference dataset, CorefUD (Nedoluzhko et al., 2022), we conduct an investigation into the t…
coreference-resolutionCoreference ResolutionMultilingual Coreference Resolution in Multiparty Dialogue
Existing multiparty dialogue datasets for entity coreference resolution are nascent, and many challenges are still unaddressed. We create a large-scale dataset, Multilingual Multiparty Coref (MMC), for this task based on…
coreference-resolutionCoreference ResolutionData AugmentationParallel Data Helps Neural Entity Coreference Resolution
Coreference resolution is the task of finding expressions that refer to the same entity in a text. Coreference models are generally trained on monolingual annotated data but annotating coreference is expensive and challe…
coreference-resolutionCoreference ResolutionCorefUD 1.0: Coreference Meets Universal Dependencies
Recent advances in standardization for annotated language resources have led to successful large scale efforts, such as the Universal Dependencies (UD) project for multilingual syntactically annotated data. By comparison…
coreference-resolutionCoreference Resolutionnamed-entity-recognitionNamed Entity Recognition+1