paper-with-me

홈 › Papers

Knowing When Not to Answer: Lightweight KB-Aligned OOD Detection for Safe RAG

2025-08-04 · Ilias Triantafyllopoulos, Renyi Qu, Salvatore Giorgi, Brenda Curtis, Lyle H. Ungar, João Sedoc arxiv

Retrieval-Augmented Generation (RAG) systems are increasingly deployed in high-stakes domains, where safety depends not only on how a system answers, but also on whether a query should be answered given a knowledge base (KB). Out-of-domain (OOD) queries can cause dense retrieval to surface weakly related context and lead the generator to produce fluent but unjustified responses. We study lightweight, KB-aligned OOD detection as an always-on gate for RAG systems. Our approach applies PCA to KB embeddings and scores queries in a compact subspace selected either by explained-variance retention (EVR) or by a separability-driven t-test ranking. We evaluate geometric semantic-search rules and lightweight classifiers across 16 domains, including high-stakes COVID-19 and Substance Use KBs, and stress-test robustness using both LLM-generated attacks and an in-the-wild 4chan attack. We find that low-dimensional detectors achieve competitive OOD performance while being faster, cheaper, and more interpretable than prompted LLM-based judges. Finally, human and LLM-based evaluations show that OOD queries primarily degrade the relevance of RAG outputs, showing the need for efficient external OOD detection to maintain safe, in-scope behavior.

📄 PDF Abstract BibTeX arXiv:2508.02296

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Answering the Unanswerable Is to Err Knowingly: Analyzing and Mitigating Abstention Failures in Large Reasoning Models

2025-08-26 · Yi Liu, Xiangyu Liu, Zequn Sun, Wei Hu arxiv

Large reasoning models (LRMs) have shown remarkable progress on complex reasoning tasks. However, some questions posed to LRMs are inherently unanswerable, such as math problems lacking sufficient conditions. We find tha…

AQUA-Bench: Beyond Finding Answers to Knowing When There Are None in Audio Question Answering

2026-01-18 · Chun-Yi Kuan, Hung-yi Lee arxiv

Recent advances in audio-aware large language models have shown strong performance on audio question answering. However, existing benchmarks mainly cover answerable questions and overlook the challenge of unanswerable on…

Question Answering

Knowing When to Stop: Dynamic Context Cutoff for Large Language Models

2025-02-03 · Roy Xie, Junlin Wang, Paul Rosu, Chunyuan Deng 외

Large language models (LLMs) process entire input contexts indiscriminately, which is inefficient in cases where the information required to answer a query is localized within the context. We present dynamic context cuto…

Token Reduction

Knowing When to Answer: Adaptive Confidence Refinement for Reliable Audio-Visual Question Answering

2026-02-04 · Dinh Phu Tran, Jihoon Jeong, Saad Wazir, Seongah Kim 외 arxiv

We present a formal problem formulation for \textit{Reliable} Audio-Visual Question Answering ($\mathcal{R}$-AVQA), where we prefer abstention over answering incorrectly. While recent AVQA models have high accuracy, thei…

Audio-visual Question Answering

Knowing When to Quit: Diagnosing and Training LLMs to Abort Futile Reasoning

2026-07-31 · Xinyan Guan, Jiali Zeng, Chunlei Xin, Yaojie Lu 외 hf

Large language models generate computationally expensive yet semantically void reasoning on beyond-capability tasks, creating risks where plausible-sounding but incorrect derivations mislead users. We characterize this f…

Reinforcement Learning