paper-with-me

Papers

Factuality of Large Language Models: A Survey

2024-02-04 · Yuxia Wang, Minghan Wang, Muhammad Arslan Manzoor, Fei Liu, Georgi Georgiev, Rocktim Jyoti Das, Preslav Nakov

Large language models (LLMs), especially when instruction-tuned for chat, have become part of our daily lives, freeing people from the process of searching, extracting, and integrating information from multiple sources by offering a straightforward answer to a variety of questions in a single place. Unfortunately, in many cases, LLM responses are factually incorrect, which limits their applicability in real-world scenarios. As a result, research on evaluating and improving the factuality of LLMs has attracted a lot of attention recently. In this survey, we critically analyze existing work with the aim to identify the major challenges and their associated causes, pointing out to potential solutions for improving the factuality of LLMs, and analyzing the obstacles to automated factuality evaluation for open-ended text generation. We further offer an outlook on where future research should go.

📄 PDF Abstract BibTeX arXiv:2402.02420

Code (0)

등록된 구현이 없습니다.

Tasks

SurveyText Generation

Similar Papers 제목 키워드 기반

Survey on Factuality in Large Language Models: Knowledge, Retrieval and Domain-Specificity

2023-10-11 · Cunxiang Wang, Xiaoze Liu, Yuanhao Yue, Xiangru Tang 외

This survey addresses the crucial issue of factuality in Large Language Models (LLMs). As LLMs find applications across diverse domains, the reliability and accuracy of their outputs become vital. We define the Factualit…

RetrievalSpecificitySurvey

Retrieving Multimodal Information for Augmented Generation: A Survey

2023-03-20 · Ruochen Zhao, Hailin Chen, Weishi Wang, Fangkai Jiao 외

As Large Language Models (LLMs) become popular, there emerged an important trend of using multimodality to augment the LLMs' generation ability, which enables LLMs to better interact with the world. However, there lacks …

RetrievalSurvey

A Survey of Large Language Models Attribution

2023-11-07 · Dongfang Li, Zetian Sun, Xinshuo Hu, Zhenyu Liu 외

Open-domain generative systems have gained significant attention in the field of conversational AI (e.g., generative search engines). This paper presents a comprehensive review of the attribution mechanisms employed by t…

Survey

FELM: Benchmarking Factuality Evaluation of Large Language Models

2023-10-01 · NeurIPS 2023 11 · Shiqi Chen, Yiran Zhao, Jinghan Zhang, I-Chun Chern 외

Assessing factuality of text generated by large language models (LLMs) is an emerging yet crucial research area, aimed at alerting users to potential errors and guiding the development of more reliable LLMs. Nonetheless,…

BenchmarkingMathRetrievalWorld Knowledge

A Survey on Multimodal Disinformation Detection

2021-03-13 · COLING 2022 10 · Firoj Alam, Stefano Cresci, Tanmoy Chakraborty, Fabrizio Silvestri 외

Recent years have witnessed the proliferation of offensive content online such as fake news, propaganda, misinformation, and disinformation. While initially this was mostly about textual content, over time images and vid…

MisinformationSurvey