MultiFC: A Real-World Multi-Domain Dataset for Evidence-Based Fact Checking of Claims
We contribute the largest publicly available dataset of naturally occurring factual claims for the purpose of automatic claim verification. It is collected from 26 fact checking websites in English, paired with textual sources and rich metadata, and labelled for veracity by human expert journalists. We present an in-depth analysis of the dataset, highlighting characteristics and challenges. Further, we present results for automatic veracity prediction, both with established baselines and with a novel method for joint ranking of evidence pages and predicting veracity that outperforms all baselines. Significant performance increases are achieved by encoding evidence, and by modelling metadata. Our best-performing model achieves a Macro F1 of 49.2%, showing that this is a challenging testbed for claim veracity prediction.
Code (0)
등록된 구현이 없습니다.
Tasks
Claim VerificationFact CheckingSimilar Papers 제목 키워드 기반
Implicit Temporal Reasoning for Evidence-Based Fact-Checking
Leveraging contextual knowledge has become standard practice in automated claim verification, yet the impact of temporal reasoning has been largely overlooked. Our study demonstrates that time positively influences the c…
Claim VerificationFact CheckingA Cognitively Grounded Bayesian Framework for Misinformation Susceptibility
In this (work in progress) paper, we present Bounded Pragmatic Listener (or BPL), a cognitively grounded Bayesian framework for modelling susceptibility to information disorder. BPL extends Rational Speech Act theory wit…
Semantic Representation and Inference for NLP
Semantic representation and inference is essential for Natural Language Processing (NLP). The state of the art for semantic representation and inference is deep learning, and particularly Recurrent Neural Networks (RNNs)…
Claim VerificationDeep LearningFact CheckingKnowledge Graphs+1Domain Adaptation for Real-World Single View 3D Reconstruction
Deep learning-based object reconstruction algorithms have shown remarkable improvements over classical methods. However, supervised learning based methods perform poorly when the training data and the test data have diff…
3D ReconstructionDomain AdaptationObject ReconstructionSingle-View 3D Reconstruction+1Rethinking Blur Synthesis for Deep Real-World Image Deblurring
In this paper, we examine the problem of real-world image deblurring and take into account two key factors for improving the performance of the deep image deblurring model, namely, training data synthesis and network arc…
DeblurringImage Deblurring