paper-with-me

홈 › Papers

RARE: Retrieval-Aware Robustness Evaluation for Retrieval-Augmented Generation Systems

2025-06-01 · Yixiao Zeng, Tianyu Cao, Danqing Wang, Xinran Zhao, Zimeng Qiu, Morteza Ziyadi, Tongshuang Wu, Lei LI

Retrieval-Augmented Generation (RAG) enhances recency and factuality in answers. However, existing evaluations rarely test how well these systems cope with real-world noise, conflicting between internal and external retrieved contexts, or fast-changing facts. We introduce Retrieval-Aware Robustness Evaluation (RARE), a unified framework and large-scale benchmark that jointly stress-tests query and document perturbations over dynamic, time-sensitive corpora. One of the central features of RARE is a knowledge-graph-driven synthesis pipeline (RARE-Get) that automatically extracts single and multi-hop relations from the customized corpus and generates multi-level question sets without manual intervention. Leveraging this pipeline, we construct a dataset (RARE-Set) spanning 400 expert-level time-sensitive finance, economics, and policy documents and 48,322 questions whose distribution evolves as the underlying sources change. To quantify resilience, we formalize retrieval-conditioned robustness metrics (RARE-Met) that capture a model's ability to remain correct or recover when queries, documents, or real-world retrieval results are systematically altered. Our results show that RAG systems exhibit surprising vulnerability to perturbations, with document robustness consistently being the weakest point regardless of generator size or architecture. RAG systems consistently show lower robustness on multi-hop queries than single-hop queries across all domains.

📄 PDF Abstract BibTeX arXiv:2506.00789

Code (1)

leililab/rare 공식 구현

Tasks

RAGRetrievalRetrieval-augmented Generation

Methods 이 논문이 사용한 방법론

Refunds@Expedia|||How do I get a full refund from Expedia? “How do I get a full refund from Expedia? How do I get a full refund from Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Quick Help &…
Linear Layer A Linear Layer is a projection $\mathbf{XW + b}$.
BPE Byte Pair Encoding, or BPE, is a subword segmentation algorithm that encodes rare and unknown words as sequences of subword units. The intuition is that various word…
Attention Dropout Attention Dropout is a type of dropout used in attention-based architectures, where elements are randomly dropped out of the…
Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…
WordPiece 설명 없음
BART BART is a denoising autoencoder for pretraining sequence-to-sequence models. It is trained by (1) corrupting text…
Weight Decay 설명 없음

Similar Papers 제목 키워드 기반

RARE: Redundancy-Aware Retrieval Evaluation Framework for High-Similarity Corpora

2026-04-21 · Hanjun Cho, Jay-Yoon Lee arxiv

Existing QA benchmarks typically assume distinct documents with minimal overlap, yet real-world retrieval-augmented generation (RAG) systems operate on corpora such as financial reports, legal codes, and patents, where i…

MADTempo: An Interactive System for Multi-Event Temporal Video Retrieval with Query Augmentation

2025-12-15 · Huu-An Vu, Van-Khanh Mai, Trong-Tam Nguyen, Quang-Duc Dam 외 arxiv

The rapid expansion of video content across online platforms has accelerated the need for retrieval systems capable of understanding not only isolated visual moments but also the temporal structure of complex events. Exi…

Visual GroundingVideo Retrieval

TABi: Type-Aware Bi-Encoders for Open-Domain Entity Retrieval

2022-04-18 · Findings (ACL) 2022 5 · Megan Leszczynski, Daniel Y. Fu, Mayee F. Chen, Christopher Ré

Entity retrieval--retrieving information about entity mentions in a query--is a key step in open-domain tasks, such as question answering or fact checking. However, state-of-the-art entity retrievers struggle to retrieve…

Entity RetrievalFact CheckingQuestion AnsweringRetrieval+1

Open-Attribute Person Retrieval: Finding People Through Distinctive and Novel Attributes

2025-08-02 · Minjeong Park, Hongbeen Park, Sangwon Lee, Jinkyu Kim arxiv

Person retrieval in surveillance videos often depends on attributes described by witnesses or operators. However, the most useful cues in practice are not always common appearance descriptions (e.g., gender, clothing col…

Person Retrieval

SearchAD: Large-Scale Rare Image Retrieval Dataset for Autonomous Driving

2026-04-09 · Felix Embacher, Jonas Uhrig, Marius Cordts, Markus Enzweiler arxiv

Retrieving rare and safety-critical driving scenarios from large-scale datasets is essential for building robust autonomous driving (AD) systems. As dataset sizes continue to grow, the key challenge shifts from collectin…

Autonomous DrivingFew-Shot LearningImage Retrieval