Evaluating anaphora and coreference resolution to improve automatic keyphrase extraction
In this paper we analyze the effectiveness of using linguistic knowledge from coreference and anaphora resolution for improving the performance for supervised keyphrase extraction. In order to verify the impact of these features, we define a baseline keyphrase extraction system and evaluate its performance on a standard dataset using different machine learning algorithms. Then, we consider new sets of features by adding combinations of the linguistic features we propose and we evaluate the new performance of the system. We also use anaphora and coreference resolution to transform the documents, trying to simulate the cohesion process performed by the human mind. We found that our approach has a slightly positive impact on the performance of automatic keyphrase extraction, in particular when considering the ranking of the results.
Code (0)
등록된 구현이 없습니다.
Tasks
Clusteringcoreference-resolutionCoreference ResolutionKeyphrase ExtractionLanguage ModelingLanguage ModellingSimilar Papers 제목 키워드 기반
CREAD: Combined Resolution of Ellipses and Anaphora in Dialogues
Anaphora and ellipses are two common phenomena in dialogues. Without resolving referring expressions and information omission, dialogue systems may fail to generate consistent and coherent responses. Traditionally, anaph…
coreference-resolutionCoreference ResolutionDialogue UnderstandingThe Universal Anaphora Scorer
The aim of the Universal Anaphora initiative is to push forward the state of the art in anaphora and anaphora resolution by expanding the aspects of anaphoric interpretation which are or can be reliably annotated in anap…
BERT-based Cohesion Analysis of Japanese Texts
The meaning of natural language text is supported by cohesion among various kinds of entities, including coreference relations, predicate-argument structures, and bridging anaphora relations. However, predicate-argument …
coreference-resolutionCoreference ResolutionThe Extended DIRNDL Corpus as a Resource for Coreference and Bridging Resolution
DIRNDL is a spoken and written corpus based on German radio news, which features coreference and information-status annotation (including bridging anaphora and their antecedents), as well as prosodic information. We have…
coreference-resolutionCoreference ResolutionAnaphora Resolution in Dialogue: System Description (CODI-CRAC 2022 Shared Task)
We describe three models submitted for the CODI-CRAC 2022 shared task. To perform identity anaphora resolution, we test several combinations of the incremental clustering approach based on the Workspace Coreference Syste…
ClusteringMulti-Task Learning