paper-with-me

홈 › Papers

RAG in the Wild: On the (In)effectiveness of LLMs with Mixture-of-Knowledge Retrieval Augmentation

2025-07-26 · Ran Xu, Yuchen Zhuang, Yue Yu, Haoyu Wang, Wenqi Shi, Carl Yang arxiv

Retrieval-augmented generation (RAG) enhances large language models (LLMs) by integrating external knowledge retrieved at inference time. While RAG demonstrates strong performance on benchmarks largely derived from general-domain corpora like Wikipedia, its effectiveness under realistic, diverse retrieval scenarios remains underexplored. We evaluated RAG systems using MassiveDS, a large-scale datastore with mixture of knowledge, and identified critical limitations: retrieval mainly benefits smaller models, rerankers add minimal value, and no single retrieval source consistently excels. Moreover, current LLMs struggle to route queries across heterogeneous knowledge sources. These findings highlight the need for adaptive retrieval strategies before deploying RAG in real-world settings. Our code and data can be found at https://github.com/ritaranx/RAG_in_the_Wild.

📄 PDF Abstract BibTeX arXiv:2507.20059

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Unveiling and Consulting Core Experts in Retrieval-Augmented MoE-based LLMs

2024-10-20 · Xin Zhou, Ping Nie, Yiwen Guo, Haojie Wei 외

Retrieval-Augmented Generation (RAG) significantly improved the ability of Large Language Models (LLMs) to solve knowledge-intensive tasks. While existing research seeks to enhance RAG performance by retrieving higher-qu…

RAGRetrievalRetrieval-augmented Generation

WildHallucinations: Evaluating Long-form Factuality in LLMs with Real-World Entity Queries

2024-07-24 · Wenting Zhao, Tanya Goyal, Yu Ying Chiu, Liwei Jiang 외

While hallucinations of large language models (LLMs) prevail as a major challenge, existing evaluation benchmarks on factuality do not cover the diverse domains of knowledge that the real-world users of LLMs seek informa…

ChatbotFormHallucinationRetrieval

Accommodate Knowledge Conflicts in Retrieval-augmented LLMs: Towards Reliable Response Generation in the Wild

2025-04-17 · Jiatai Wang, Zhiwei Xu, Di Jin, Xuewen Yang 외

The proliferation of large language models (LLMs) has significantly advanced information retrieval systems, particularly in response generation (RG). Unfortunately, LLMs often face knowledge conflicts between internal me…

Decision MakingInformation RetrievalMisinformationNavigate+6

WildDESED: An LLM-Powered Dataset for Wild Domestic Environment Sound Event Detection System

2024-07-04 · Yang Xiao, Rohan Kumar Das

This work aims to advance sound event detection (SED) research by presenting a new large language model (LLM)-powered dataset namely wild domestic environment sound event detection (WildDESED). It is crafted as an extens…

Event DetectionLanguage ModelingLanguage ModellingLarge Language Model+1

HaluEval-Wild: Evaluating Hallucinations of Language Models in the Wild

2024-03-07 · Zhiying Zhu, Yiming Yang, Zhiqing Sun

Hallucinations pose a significant challenge to the reliability of large language models (LLMs) in critical domains. Recent benchmarks designed to assess LLM hallucinations within conventional NLP tasks, such as knowledge…

HallucinationQuestion AnsweringRAGRetrieval+1