Quotations, Coreference Resolution, and Sentiment Annotations in Croatian News Articles: An Exploratory Study
This paper presents a corpus annotated for the task of direct-speech extraction in Croatian. The paper focuses on the annotation of the quotation, co-reference resolution, and sentiment annotation in SETimes news corpus in Croatian and on the analysis of its language-specific differences compared to English. From this, a list of the phenomena that require special attention when performing these annotations is derived. The generated corpus with quotation features annotations can be used for multiple tasks in the field of Natural Language Processing.
Code (0)
등록된 구현이 없습니다.
Tasks
Articlescoreference-resolutionCoreference ResolutionSpeech ExtractionSimilar Papers 제목 키워드 기반
The Project Dialogism Novel Corpus: A Dataset for Quotation Attribution in Literary Texts
We present the Project Dialogism Novel Corpus, or PDNC, an annotated dataset of quotations for English literary texts. PDNC contains annotations for 35,978 quotations across 22 full-length novels, and is by an order of m…
Referring ExpressionImproving Automatic Quotation Attribution in Literary Novels
Current models for quotation attribution in literary novels assume varying levels of available information in their training and test data, which poses a challenge for in-the-wild inference. Here, we approach quotation a…
coreference-resolutionCoreference ResolutionResolving Entity Coreference in Croatian with a Constrained Mention-Pair Model
It’s absolutely divine! Can fine-grained sentiment analysis benefit from coreference resolution?
While it has been claimed that anaphora or coreference resolution plays an important role in opinion mining, it is not clear to what extent coreference resolution actually boosts performance, if at all. In this paper, we…
Aspect-Based Sentiment AnalysisAspect-Based Sentiment Analysis (ABSA)coreference-resolutionCoreference Resolution+2Marmara Turkish Coreference Corpus and Coreference Resolution Baseline
We describe the Marmara Turkish Coreference Corpus, which is an annotation of the whole METU-Sabanci Turkish Treebank with mentions and coreference chains. Collecting eight or more independent annotations for each docume…
coreference-resolutionCoreference Resolution