paper-with-me

홈 › Papers

An Audit on the Perspectives and Challenges of Hallucinations in NLP

2024-04-11 · Pranav Narayanan Venkit, Tatiana Chakravorti, Vipul Gupta, Heidi Biggs, Mukund Srinath, Koustava Goswami, Sarah Rajtmajer, Shomir Wilson

We audit how hallucination in large language models (LLMs) is characterized in peer-reviewed literature, using a critical examination of 103 publications across NLP research. Through the examination of the literature, we identify a lack of agreement with the term `hallucination' in the field of NLP. Additionally, to compliment our audit, we conduct a survey with 171 practitioners from the field of NLP and AI to capture varying perspectives on hallucination. Our analysis calls for the necessity of explicit definitions and frameworks outlining hallucination within NLP, highlighting potential challenges, and our survey inputs provide a thematic understanding of the influence and ramifications of hallucination in society.

📄 PDF Abstract BibTeX arXiv:2404.07461

Code (0)

등록된 구현이 없습니다.

Tasks

HallucinationSurvey

Similar Papers 제목 키워드 기반

Trouble du contr\^ole de la parole int\'erieure : cas des hallucinations auditives verbales (Inner speech monitoring deficit : a study of auditory verbal hallucinations) [in French]

2012-06-01 · JEPTALNRECITAL 2012 6 · Lucile Rapin, Marion Dohen, H{\'e}l{\`e}ne L{\oe}venbruck, Mircea Polosan 외

OmniDPO: A Preference Optimization Framework to Address Omni-Modal Hallucination

2025-08-31 · Junzhe Chen, Tianshu Zhang, Shiyu Huang, Yuwei Niu 외 arxiv

Recently, Omni-modal large language models (OLLMs) have sparked a new wave of research, achieving impressive results in tasks such as audio-video understanding and real-time environment perception. However, hallucination…

EGOILLUSION: Benchmarking Hallucinations in Egocentric Video Understanding

2025-08-18 · Ashish Seth, Utkarsh Tyagi, Ramaneswaran Selvakumar, Nishit Anand 외 arxiv

Multimodal Large Language Models (MLLMs) have demonstrated remarkable performance in complex multimodal tasks. While MLLMs excel at visual perception and reasoning in third-person and egocentric videos, they are prone to…

Reefknot: A Comprehensive Benchmark for Relation Hallucination Evaluation, Analysis and Mitigation in Multimodal Large Language Models

2024-08-18 · Kening Zheng, Junkai Chen, Yibo Yan, Xin Zou 외

Hallucination issues continue to affect multimodal large language models (MLLMs), with existing research mainly addressing object-level or attribute-level hallucinations, neglecting the more complex relation hallucinatio…

AttributeHallucinationHallucination EvaluationRelation

A Blueprint for Auditing Generative AI

2024-07-07 · Jakob Mokander, Justin Curl, Mihir Kshirsagar

The widespread use of generative AI systems is coupled with significant ethical and social challenges. As a result, policymakers, academic researchers, and social advocacy groups have all called for such systems to be au…