The Perils & Promises of Fact-checking with Large Language Models
Automated fact-checking, using machine learning to verify claims, has grown vital as misinformation spreads beyond human fact-checking capacity. Large Language Models (LLMs) like GPT-4 are increasingly trusted to write academic papers, lawsuits, and news articles and to verify information, emphasizing their role in discerning truth from falsehood and the importance of being able to verify their outputs. Understanding the capacities and limitations of LLMs in fact-checking tasks is therefore essential for ensuring the health of our information ecosystem. Here, we evaluate the use of LLM agents in fact-checking by having them phrase queries, retrieve contextual data, and make decisions. Importantly, in our framework, agents explain their reasoning and cite the relevant sources from the retrieved context. Our results show the enhanced prowess of LLMs when equipped with contextual information. GPT-4 outperforms GPT-3, but accuracy varies based on query language and claim veracity. While LLMs show promise in fact-checking, caution is essential due to inconsistent accuracy. Our investigation calls for further research, fostering a deeper comprehension of when agents succeed and when they fail.
Code (0)
등록된 구현이 없습니다.
Tasks
ArticlesFact CheckingMisinformationMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Here be dragons? The perils and promises of inter-resource lexical-semantic mapping
Community Moderation and the New Epistemology of Fact Checking on Social Media
Social media platforms have traditionally relied on internal moderation teams and partnerships with independent fact-checking organizations to identify and flag misleading content. Recently, however, platforms including …
Fact CheckingMisinformationGenerative Large Language Models in Automated Fact-Checking: A Survey
The dissemination of false information on online platforms presents a serious societal challenge. While manual fact-checking remains crucial, Large Language Models (LLMs) offer promising opportunities to support fact-che…
Fact CheckingSurveySelf-Checker: Plug-and-Play Modules for Fact-Checking with Large Language Models
Fact-checking is an essential task in NLP that is commonly utilized for validating the factual accuracy of claims. Prior work has mainly focused on fine-tuning pre-trained languages models on specific datasets, which can…
Fact CheckingIn-Context LearningEvidence-based Interpretable Open-domain Fact-checking with Large Language Models
Universal fact-checking systems for real-world claims face significant challenges in gathering valid and sufficient real-time evidence and making reasoned decisions. In this work, we introduce the Open-domain Explainable…
Fact Checkingvalid