paper-with-me

홈 › Papers

WDV: A Broad Data Verbalisation Dataset Built from Wikidata

2022-05-05 · Gabriel Amaral, Odinaldo Rodrigues, Elena Simperl

Data verbalisation is a task of great importance in the current field of natural language processing, as there is great benefit in the transformation of our abundant structured and semi-structured data into human-readable formats. Verbalising Knowledge Graph (KG) data focuses on converting interconnected triple-based claims, formed of subject, predicate, and object, into text. Although KG verbalisation datasets exist for some KGs, there are still gaps in their fitness for use in many scenarios. This is especially true for Wikidata, where available datasets either loosely couple claim sets with textual information or heavily focus on predicates around biographies, cities, and countries. To address these gaps, we propose WDV, a large KG claim verbalisation dataset built from Wikidata, with a tight coupling between triples and text, covering a wide variety of entities and predicates. We also evaluate the quality of our verbalisations through a reusable workflow for measuring human-centred fluency and adequacy scores. Our data and code are openly available in the hopes of furthering research towards KG verbalisation.

📄 PDF Abstract BibTeX arXiv:2205.02627

Code (1)

gabrielmaia7/wdv 공식 구현 pytorch

Similar Papers 제목 키워드 기반

Assessing the quality of sources in Wikidata across languages: a hybrid approach

2021-09-20 · Gabriel Amaral, Alessandro Piscopo, Lucie-Aimée Kaffee, Odinaldo Rodrigues 외

Wikidata is one of the most important sources of structured data on the web, built by a worldwide community of volunteers. As a secondary source, its contents must be backed by credible references; this is particularly i…

Descriptive

Does Wikidata Support Analogical Reasoning?

2022-10-02 · Filip Ilievski, Jay Pujara, Kartik Shenoy

Analogical reasoning methods have been built over various resources, including commonsense knowledge bases, lexical resources, language models, or their combination. While the wide coverage of knowledge about entities an…

Learning to Recommend Items to Wikidata Editors

2021-07-13 · Kholoud Alghamdi, Miaojing Shi, Elena Simperl

Wikidata is an open knowledge graph built by a global community of volunteers. As it advances in scale, it faces substantial challenges around editor engagement. These challenges are in terms of both attracting new edito…

Collaborative FilteringRecommendation Systems

ProVe: A Pipeline for Automated Provenance Verification of Knowledge Graphs against Textual Sources

2022-10-26 · Gabriel Amaral, Odinaldo Rodrigues, Elena Simperl

Knowledge Graphs are repositories of information that gather data from a multitude of domains and sources in the form of semantic triples, serving as a source of structured data for various crucial applications in the mo…

Binary ClassificationClaim VerificationKnowledge GraphsSentence

Enriching Wikidata with Linked Open Data

2022-07-01 · Bohui Zhang, Filip Ilievski, Pedro Szekely

Large public knowledge graphs, like Wikidata, contain billions of statements about tens of millions of entities, thus inspiring various use cases to exploit such knowledge graphs. However, practice shows that much of the…

Entity AlignmentKnowledge Graphs