Knowledge Extraction From Texts Based on Wikidata
This paper presents an effort within our company of developing knowledge extraction pipeline for English, which can be further used for constructing an entreprise-specific knowledge base. We present a system consisting of entity detection and linking, coreference resolution, and relation extraction based on the Wikidata schema. We highlight existing challenges of knowledge extraction by evaluating the deployed pipeline on real-world data. We also make available a database, which can serve as a new resource for sentential relation extraction, and we underline the importance of having balanced data for training classification models.
Code (1)
Tasks
coreference-resolutionCoreference ResolutionRelationRelation ExtractionSimilar Papers 제목 키워드 기반
Scholarly Wikidata: Population and Exploration of Conference Data in Wikidata using LLMs
Several initiatives have been undertaken to conceptually model the domain of scholarly data using ontologies and to create respective Knowledge Graphs. Yet, the full potential seems unleashed, as automated means for auto…
Knowledge GraphsDaMuEL: A Large Multilingual Dataset for Entity Linking
We present DaMuEL, a large Multilingual Dataset for Entity Linking containing data in 53 languages. DaMuEL consists of two components: a knowledge base that contains language-agnostic information about entities, includin…
Entity LinkingTemplate-based multilingual football reports generation using Wikidata as a knowledge base
This paper presents a new version of a football reports generation system called PASS. The original version generated Dutch text and relied on a limited hand-crafted knowledge base. We describe how, in a short amount of …
Machine TranslationText GenerationTranslationTowards a Brazilian History Knowledge Graph
This short paper describes the first steps in a project to construct a knowledge graph for Brazilian history based on the Brazilian Dictionary of Historical Biographies (DHBB) and Wikipedia/Wikidata. We contend that larg…
Enriching Knowledge Bases with Counting Quantifiers
Information extraction traditionally focuses on extracting relations between identifiable entities, such as <Monterey, locatedIn, California>. Yet, texts often also contain Counting information, stating that a subject is…