paper-with-me

홈 › Papers

LANCET: Neural Intervention via Structural Entropy for Mitigating Faithfulness Hallucinations in LLMs

2026-01-04 · Chenxu Wang, Chaozhuo Li, Pengbo Wang, Litian Zhang, Songyang Liu, Ji Qi, Jiahui Hu, Yushan Cai, Hao Zhao, Rui Pu arxiv

Large Language Models have revolutionized information processing, yet their reliability is severely compromised by faithfulness hallucinations. While current approaches attempt to mitigate this issue through node-level adjustments or coarse suppression, they often overlook the distributed nature of neural information, leading to imprecise interventions. Recognizing that hallucinations propagate through specific forward transmission pathways like an infection, we aim to surgically block this flow using precise structural analysis. To leverage this, we propose Lancet, a novel framework that achieves precise neural intervention by leveraging structural entropy and hallucination difference ratios. Lancet first locates hallucination-prone neurons via gradient-driven contrastive analysis, then maps their propagation pathways by minimizing structural entropy, and finally implements a hierarchical intervention strategy that preserves general model capabilities. Comprehensive evaluations across hallucination benchmark datasets demonstrate that Lancet significantly outperforms state-of-the-art methods, validating the effectiveness of our surgical approach to neural intervention.

📄 PDF Abstract BibTeX arXiv:2601.01401

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Faithfulness as Information Flow: Evaluating and Training Faithful Chain-of-Thought Reasoning

2026-05-22 · Jinghan Jia, Joe Benton, Eric Easley arxiv

Chain-of-thought (CoT) reasoning is useful for monitoring language models only when the reasoning trace faithfully reflects the computation that produces the final answer. However, models can rely on prompt-to-answer sho…

Code Repair

Dismantling Pathological Shortcuts: A Causal Framework for Faithful LVLM Decoding

2026-06-25 · Liu Yu, Can Chen, Ping Kuang, Zhikun Feng 외 arxiv

Large Vision-Language Models (LVLMs) exhibit sophisticated reasoning but remain susceptible to object hallucination. Deviating from the prevailing attention intensity assumption, we reveal a deeper dynamic structural mis…

Visual Grounding

RFEval: Benchmarking Reasoning Faithfulness under Counterfactual Reasoning Intervention in Large Reasoning Models

2026-02-19 · Yunseok Han, Yejoon Lee, Jaeyoung Do arxiv

Large Reasoning Models (LRMs) exhibit strong performance, yet often produce rationales that sound plausible but fail to reflect their true decision process, undermining reliability and trust. We introduce a formal framew…

Faithfulness Serum: Mitigating the Faithfulness Gap in Textual Explanations of LLM Decisions via Attribution Guidance

2026-04-15 · Bar Alon, Itamar Zimerman, Lior Wolf arxiv

Large language models (LLMs) achieve strong performance and have revolutionized NLP, but their lack of explainability keeps them treated as black boxes, limiting their use in domains that demand transparency and trust. A…

Explanation Generation

Causal Layering via Conditional Entropy

2024-01-19 · Itai Feigenbaum, Devansh Arpit, Huan Wang, Shelby Heinecke 외

Causal discovery aims to recover information about an unobserved causal graph from the observable data it generates. Layerings are orderings of the variables which place causes before effects. In this paper, we provide w…

Causal Discovery