KnowFi: Knowledge Extraction from Long Fictional Texts
Knowledge base construction has recently been extended to fictional domains like multi-volume novels and TV/movie series, aiming to support explorative queries for fans and sub-culture studies by humanities researchers. This task involves the extraction of relations between entities. State-of-the-art methods are geared for short input texts and basic relations, but fictional domains require tapping very long texts and need to cope with non-standard relations where distant supervision becomes sparse. This work addresses these challenges by a novel method, called KnowFi, that combines BERT-enhanced neural learning with judicious selection and aggregation of text passages. Experiments with several fictional domains demonstrate the gains that KnowFi achieves over the best prior methods for neural relation extraction.
Code (0)
등록된 구현이 없습니다.
Tasks
Cultural Vocal Bursts Intensity PredictionKnowledge Base ConstructionRelation ExtractionSimilar Papers 제목 키워드 기반
ENTYFI: A System for Fine-grained Entity Typing in Fictional Texts
Fiction and fantasy are archetypes of long-tail domains that lack suitable NLP methodologies and tools. We present ENTYFI, a web-based system for fine-grained typing of entity mentions in fictional texts. It builds on 20…
DiversityEntity TypingVocal Bursts Type PredictionTiFi: Taxonomy Induction for Fictional Domains [Extended version]
Taxonomies are important building blocks of structured knowledge bases, and their construction from text sources and Wikipedia has received much attention. In this paper we focus on the construction of taxonomies for fic…
FantasyCoref: Coreference Resolution on Fantasy Literature Through Omniscient Writer’s Point of View
This paper presents a new corpus and annotation guideline for a novel coreference resolution task on fictional texts, and analyzes its unique characteristics. FantasyCoref contains 211 stories of Grimms’ Fairy Tales and …
coreference-resolutionCoreference ResolutionSentiment Analysis with R: Natural Language Processing for Semi-Automated Assessments of Qualitative Data
Sentiment analysis is a sub-discipline in the field of natural language processing and computational linguistics and can be used for automated or semi-automated analyses of text documents. One of the aims of these analys…
Sentiment AnalysisLevels of Non-Fictionality in Fictional Texts
The annotation and automatic recognition of non-fictional discourse within a text is an important, yet unresolved task in literary research. While non-fictional passages can consist of several clauses or sentences, we ar…