paper-with-me

홈 › Papers

Retrieve Anything To Augment Large Language Models

2023-10-11 · Peitian Zhang, Shitao Xiao, Zheng Liu, Zhicheng Dou, Jian-Yun Nie

Large language models (LLMs) face significant challenges stemming from their inherent limitations in knowledge, memory, alignment, and action. These challenges cannot be addressed by LLMs alone, but should rely on assistance from the external world, such as knowledge base, memory store, demonstration examples, and tools. Retrieval augmentation stands as a vital mechanism for bridging the gap between LLMs and the external assistance. However, conventional methods encounter two pressing issues. On the one hand, the general-purpose retrievers are not properly optimized for the retrieval augmentation of LLMs. On the other hand, the task-specific retrievers lack the required versatility, hindering their performance across the diverse retrieval augmentation scenarios. In this work, we present a novel approach, the LLM-Embedder, which comprehensively supports the diverse retrieval augmentation needs of LLMs with one unified embedding model. Training such a unified model is non-trivial, as various retrieval tasks aim to capture distinct semantic relationships, often subject to mutual interference. To address this challenge, we systematically optimize our training methodology. This includes reward formulation based on LLMs' feedback, the stabilization of knowledge distillation, multi-task fine-tuning with explicit instructions, and homogeneous in-batch negative sampling. These optimization strategies contribute to the outstanding empirical performance of the LLM-Embedder. Notably, it yields remarkable enhancements in retrieval augmentation for LLMs, surpassing both general-purpose and task-specific retrievers in various evaluation scenarios. Our checkpoint and source code are publicly available at https://github.com/FlagOpen/FlagEmbedding.

📄 PDF Abstract BibTeX arXiv:2310.07554

Code (1)

flagopen/flagembedding 공식 구현 pytorch

Tasks

Knowledge DistillationRetrieval

Similar Papers 제목 키워드 기반

Anything2Skill: Compiling External Knowledge into Reusable Skills for Agents

2026-06-08 · Qianjun Pan, Yutao Yang, Junsong Li, Jie Zhou 외 arxiv

Retrieval-augmented generation (RAG) enables agents to access external knowledge at inference time, but it primarily retrieves fragmented declarative evidence, leaving agents to repeatedly infer task procedures from pass…

Can Retriever-Augmented Language Models Reason? The Blame Game Between the Retriever and the Language Model

2022-12-18 · Parishad BehnamGhader, Santiago Miret, Siva Reddy

Augmenting pretrained language models with retrievers has shown promise in effectively solving common NLP problems, such as language modeling and question answering. In this paper, we evaluate the strengths and weaknesse…

Language ModelingLanguage ModellingQuestion AnsweringRetrieval

Composition Vision-Language Understanding via Segment and Depth Anything Model

2024-06-07 · Mingxiao Huo, Pengliang Ji, Haotian Lin, Junchen Liu 외

We introduce a pioneering unified library that leverages depth anything, segment anything models to augment neural comprehension in language-vision model zero-shot understanding. This library synergizes the capabilities …

Question AnsweringVisual Question Answering (VQA)

Augmentation-Adapted Retriever Improves Generalization of Language Models as Generic Plug-In

2023-05-27 · Zichun Yu, Chenyan Xiong, Shi Yu, Zhiyuan Liu

Retrieval augmentation can aid language models (LMs) in knowledge-intensive tasks by supplying them with external information. Prior works on retrieval augmentation usually jointly fine-tune the retriever and the LM, mak…

MMLURetrievalZero-shot Generalization

Caption Anything: Interactive Image Description with Diverse Multimodal Controls

2023-05-04 · Teng Wang, Jinrui Zhang, Junjie Fei, Hao Zheng 외

Controllable image captioning is an emerging multimodal topic that aims to describe the image with natural language following human purpose, $\textit{e.g.}$, looking at the specified regions or telling in a particular te…

controllable image captioningImage CaptioningImage DescriptionInstruction Following