paper-with-me

홈 › Papers

ACL-Verbatim: hallucination-free question answering for research

2026-05-20 · Gábor Recski, Szilveszter Tóth, Nadia Verdha, István Boros, Ádám Kovács arxiv

Academic researchers need efficient and reliable methods for collecting high-quality information from trusted sources, but modern tools for AI-assisted research still suffer from the tendency of Large Language Models (LLMs) to produce factually inaccurate or nonsensical output, commonly referred to as hallucinations. We apply the extractive question answering system VerbatimRAG to research papers in the ACL Anthology, directly mapping user queries to verbatim text spans in retrieved documents. We contribute a novel ground truth dataset for the task of mapping user queries to relevant text spans in research papers, and use it to train and evaluate a variety of extractive models. Human annotation is performed by NLP researchers and is based on synthetic user queries generated using a custom pipeline based on the ScIRGen methodology, paired with chunks of research papers retrieved by VerbatimRAG. On this benchmark, a 150M-parameter ModernBERT token classifier trained on silver supervision from our pipeline achieves the best word-level F1 (53.6), ahead of the strongest evaluated LLM extractor (48.7).

📄 PDF Abstract BibTeX arXiv:2605.21102

Code (0)

등록된 구현이 없습니다.

Tasks

Question Answering

Similar Papers 제목 키워드 기반

Peering into the Mind of Language Models: An Approach for Attribution in Contextual Question Answering

2024-05-28 · Anirudh Phukan, Shwetha Somasundaram, Apoorv Saxena, Koustava Goswami 외

With the enhancement in the field of generative artificial intelligence (AI), contextual question answering has become extremely relevant. Attributing model generations to the input source document is essential to ensure…

Question Answering

HaluMem: Evaluating Hallucinations in Memory Systems of Agents

2025-11-05 · Ding Chen, Simin Niu, Kehang Li, Peng Liu 외 arxiv

Memory systems are key components that enable AI systems such as LLMs and AI agents to achieve long-term learning and sustained interaction. However, during memory storage and retrieval, these systems frequently exhibit …

Question Answering

KSHSeek: Data-Driven Approaches to Mitigating and Detecting Knowledge-Shortcut Hallucinations in Generative Models

2025-03-25 · Zhiwei Wang, Zhongxin Liu, Ying Li, Hongyu Sun 외

The emergence of large language models (LLMs) has significantly advanced the development of natural language processing (NLP), especially in text generation tasks like question answering. However, model hallucinations re…

HallucinationQuestion AnsweringText Generation

DelucionQA: Detecting Hallucinations in Domain-specific Question Answering

2023-12-08 · Mobashir Sadat, Zhengyu Zhou, Lukas Lange, Jun Araki 외

Hallucination is a well-known phenomenon in text generated by large language models (LLMs). The existence of hallucinatory responses is found in almost all application scenarios e.g., summarization, question-answering (Q…

HallucinationInformation RetrievalQuestion AnsweringRetrieval

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering

2024-11-19 · Aryan Keluskar, Amrita Bhattacharjee, Huan Liu

Ambiguity in natural language poses significant challenges to Large Language Models (LLMs) used for open-domain question answering. LLMs often struggle with the inherent uncertainties of human communication, leading to m…

Fact CheckingOpen-Domain Question AnsweringQuestion AnsweringSentiment Analysis