paper-with-me

홈 › Papers

Faithfulness-QA: A Counterfactual Entity Substitution Dataset for Training Context-Faithful RAG Models

2026-04-28 · Li Ju, Junzhe Wang, Qi Zhang arxiv

Retrieval-Augmented Generation (RAG) models frequently produce answers grounded in parametric memory rather than the retrieved context, undermining the core promise of retrieval augmentation. A fundamental obstacle to fixing this unfaithfulness is the lack of training data that explicitly requires models to prefer context over internal knowledge. We introduce Faithfulness-QA, a large-scale dataset of 99,094 samples constructed through counterfactual entity substitution. Starting from two established extractive QA benchmarks--SQuAD and TriviaQA--we automatically identify answer-bearing named entities in each context, replace them with type-consistent alternatives drawn from a curated bank of 76,953 entities, and thereby manufacture controlled knowledge conflicts between context and parametric memory. Rigorous quality filtering ensures 100% pass rates across four automated checks on random 200-sample audits. We release the full dataset, the construction pipeline, and a typed entity bank covering eight named entity categories. Faithfulness-QA is designed as a training resource for attention-based faithfulness objectives and as an evaluation benchmark for measuring context-grounding behavior in RAG systems. Data and code are available at https://github.com/qzhangFDU/faithfulness-qa-dataset.

📄 PDF Abstract BibTeX arXiv:2604.25313

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Does Your Model Classify Entities Reasonably? Diagnosing and Mitigating Spurious Correlations in Entity Typing

2022-05-25 · Nan Xu, Fei Wang, Bangzheng Li, Mingtao Dong 외

Entity typing aims at predicting one or more words that describe the type(s) of a specific mention in a sentence. Due to shortcuts from surface patterns to annotated entity labels and biased training, existing entity typ…

counterfactualData AugmentationEntity TypingSentence

Counterfactual Simulation Training for Chain-of-Thought Faithfulness

2026-02-24 · Peter Hase, Christopher Potts arxiv

Inspecting Chain-of-Thought reasoning is among the most common means of understanding why an LLM produced its output. But well-known problems with CoT faithfulness severely limit what insights can be gained from this pra…

Diffusion Counterfactual Generation with Semantic Abduction

2025-06-09 · Rajat Rasal, Avinash Kori, Fabio De Sousa Ribeiro, Tian Xia 외

Counterfactual image generation presents significant challenges, including preserving identity, maintaining perceptual quality, and ensuring faithfulness to an underlying causal model. While existing auto-encoding framew…

counterfactualCounterfactual ReasoningImage GenerationRepresentation Learning

Counterfactual Evaluation for Explainable AI

2021-09-05 · Yingqiang Ge, Shuchang Liu, Zelong Li, Shuyuan Xu 외

While recent years have witnessed the emergence of various explainable methods in machine learning, to what degree the explanations really represent the reasoning process behind the model prediction -- namely, the faithf…

counterfactualCounterfactual Reasoning

Logical Satisfiability of Counterfactuals for Faithful Explanations in NLI

2022-01-16 · ACL ARR January 2022 1 · Anonymous

Evaluating an explanation's faithfulness is desired for many reasons such as trust, interpretability and diagnosing the sources of model's errors. In this work, which focuses on the NLI task, we introduce the methodology…

counterfactual