paper-with-me

홈 › Papers

FIDES: Faithful Inference via Deep Evidence Signals for Retrieval-Memory Conflict in RAG

2026-06-04 · Zhe Yu, Wenpeng Xing, Tiancheng Zhao, Mohan Li, Changting Lin, Meng Han arxiv

When retrieved evidence contradicts parametric memory, language models frequently ignore context and default to memorized priors -- a failure that undermines the core purpose of retrieval augmentation. Contrastive decoding amplifies the context-conditioned output to suppress parametric bias, but existing methods rest on an implicit assumption that this bias is uniform across tokens. A single global contrastive weight over-penalizes safe tokens while leaving genuinely conflicted ones insufficiently corrected. We identify token-level conflict concentration: retrieval-memory tension is sharply heterogeneous, concentrated on a small fraction of answer-critical decoding steps. This reframes contrastive decoding from how much contrast to apply to where to apply it. We propose FIDES (Faithful Inference via Deep Evidence Signals), a training-free decoder that reads three internal signals probing retrieval-memory conflict at complementary depths -- output surface, hidden representations, and prediction trajectory -- and fuses them to govern intervention strength at each decoding step. Across three benchmarks and six backbones -- four primary 7B/8B models and two scaling backbones up to 70B -- FIDES achieves the best context fidelity in all 18 settings, outperforming the strongest training-free baseline by +3 to +13 points. On the 70B scale, fidelity reaches 92-94% while F1 surges to 62-63%, demonstrating that token-level selectivity unlocks generation capability that coarse contrastive rules suppress.

📄 PDF Abstract BibTeX arXiv:2606.05644

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Decomposing and Revising What Language Models Generate

2025-08-31 · Zhichao Yan, Jiaoyan Chen, Jiapu Wang, Xiaoli Li 외 arxiv

Attribution is crucial in question answering (QA) with Large Language Models (LLMs).SOTA question decomposition-based approaches use long form answers to generate questions for retrieving related documents. However, the …

Question Answering

FiDeSR: High-Fidelity and Detail-Preserving One-Step Diffusion Super-Resolution

2026-03-03 · Aro Kim, Myeongjin Jang, Chaewon Moon, Youngjin Shin 외 arxiv

Diffusion-based approaches have recently driven remarkable progress in real-world image super-resolution (SR). However, existing methods still struggle to simultaneously preserve fine details and ensure high-fidelity rec…

Image Super-Resolution

MedRAGChecker: Claim-Level Verification for Biomedical Retrieval-Augmented Generation

2026-01-10 · Yuelyu Ji, Min Gu Kwak, Hang Zhang, Xizhi Wu 외 arxiv

Biomedical retrieval-augmented generation (RAG) can ground LLM answers in medical literature, yet long-form outputs often contain isolated unsupported or contradictory claims with safety implications. We introduce MedRAG…

Natural Language Inference

A Generative Framework for Low-Cost Result Validation of Machine Learning-as-a-Service Inference

2023-03-31 · Abhinav Kumar, Miguel A. Guirao Aguilera, Reza Tourani, Satyajayant Misra

The growing popularity of Machine Learning (ML) has led to its deployment in various sensitive domains, which has resulted in significant research focused on ML security and privacy. However, in some applications, such a…

Autonomous DrivingGenerative Adversarial NetworkTransfer Learning

Beyond Topical Similarity: Contrastive Evidence Retrieval with Interpretable Attention Alignment in RAG

2026-05-31 · Francielle Vargas, João Robiatti, Diego Alves, Lucas Pascotti Valem 외 arxiv

Ensuring factuality and interpretability in RAG remains an open and urgent problem. We introduce Contrastive Evidence Rationale Attention (CERA), the first retrieval framework to employ subjectivity-based hard negative s…

Contrastive Learning