paper-with-me

홈 › Papers

Unleashing the Power of LLMs in Dense Retrieval with Query Likelihood Modeling

2025-04-07 · Hengran Zhang, Keping Bi, Jiafeng Guo, Xiaojie Sun, Shihao Liu, Daiting Shi, Dawei Yin, Xueqi Cheng

Dense retrieval is a crucial task in Information Retrieval (IR) and is the foundation for downstream tasks such as re-ranking. Recently, large language models (LLMs) have shown compelling semantic understanding capabilities and are appealing to researchers studying dense retrieval. LLMs, as decoder-style generative models, are competent at language generation while falling short on modeling global information due to the lack of attention to tokens afterward. Inspired by the classical word-based language modeling approach for IR, i.e., the query likelihood (QL) model, we seek to sufficiently utilize LLMs' generative ability by QL maximization. However, instead of ranking documents with QL estimation, we introduce an auxiliary task of QL maximization to yield a better backbone for contrastively learning a discriminative retriever. We name our model as LLM-QL. To condense global document semantics to a single vector during QL modeling, LLM-QL has two major components, Attention Stop (AS) and Input Corruption (IC). AS stops the attention of predictive tokens to previous tokens until the ending token of the document. IC masks a portion of tokens in the input documents during prediction. Experiments on MSMARCO show that LLM-QL can achieve significantly better performance than other LLM-based retrievers and using QL estimated by LLM-QL for ranking outperforms word-based QL by a large margin.

📄 PDF Abstract BibTeX arXiv:2504.05216

Code (0)

등록된 구현이 없습니다.

Tasks

Information RetrievalLanguage ModelingLanguage ModellingRe-RankingRetrievalText Generation

Methods 이 논문이 사용한 방법론

Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…
Attention 설명 없음

Similar Papers 제목 키워드 기반

ScalingNote: Scaling up Retrievers with Large Language Models for Real-World Dense Retrieval

2024-11-24 · Suyuan Huang, Chao Zhang, Yuanyuan Wu, Haoxin Zhang 외

Dense retrieval in most industries employs dual-tower architectures to retrieve query-relevant documents. Due to online deployment requirements, existing real-world dense retrieval systems mainly enhance performance by d…

Retrieval

Question-Based Retrieval using Atomic Units for Enterprise RAG

2024-05-20 · Vatsal Raina, Mark Gales

Enterprise retrieval augmented generation (RAG) offers a highly flexible framework for combining powerful large language models (LLMs) with internal, possibly temporally changing, documents. In RAG, documents are first c…

RAGRetrievalRetrieval-augmented Generation

LLM-QE: Improving Query Expansion by Aligning Large Language Models with Ranking Preferences

2025-02-24 · Sijia Yao, Pengcheng Huang, Zhenghao Liu, Yu Gu 외

Query expansion plays a crucial role in information retrieval, which aims to bridge the semantic gap between queries and documents to improve matching performance. This paper introduces LLM-QE, a novel approach that leve…

HallucinationInformation RetrievalRetrieval

Boosting legal case retrieval by query content selection with large language models

2023-12-06 · Youchao Zhou, Heyan Huang, Zhijing Wu

Legal case retrieval, which aims to retrieve relevant cases to a given query case, benefits judgment justice and attracts increasing attention. Unlike generic retrieval queries, legal case queries are typically long and …

Retrieval

Your Dense Retriever is Secretly an Expeditious Reasoner

2025-09-27 · Yichi Zhang, Jun Bai, Zhixin Cai, Shuhan Qin 외 arxiv

Dense retrievers enhance retrieval by encoding queries and documents into continuous vectors, but they often struggle with reasoning-intensive queries. Although Large Language Models (LLMs) can reformulate queries to cap…