paper-with-me

홈 › Papers

Scaling Laws For Dense Retrieval

2024-03-27 · Yan Fang, Jingtao Zhan, Qingyao Ai, Jiaxin Mao, Weihang Su, Jia Chen, Yiqun Liu

Scaling up neural models has yielded significant advancements in a wide array of tasks, particularly in language generation. Previous studies have found that the performance of neural models frequently adheres to predictable scaling laws, correlated with factors such as training set size and model size. This insight is invaluable, especially as large-scale experiments grow increasingly resource-intensive. Yet, such scaling law has not been fully explored in dense retrieval due to the discrete nature of retrieval metrics and complex relationships between training data and model sizes in retrieval tasks. In this study, we investigate whether the performance of dense retrieval models follows the scaling law as other neural models. We propose to use contrastive log-likelihood as the evaluation metric and conduct extensive experiments with dense retrieval models implemented with different numbers of parameters and trained with different amounts of annotated data. Results indicate that, under our settings, the performance of dense retrieval models follows a precise power-law scaling related to the model size and the number of annotations. Additionally, we examine scaling with prevalent data augmentation methods to assess the impact of annotation quality, and apply the scaling law to find the best resource allocation strategy under a budget constraint. We believe that these insights will significantly contribute to understanding the scaling effect of dense retrieval models and offer meaningful guidance for future research endeavors.

📄 PDF Abstract BibTeX arXiv:2403.18684

Code (1)

jingtaozhan/drscale 공식 구현 pytorch

Tasks

Data AugmentationRetrievalText Generation

Methods 이 논문이 사용한 방법론

SET Dynamic Sparse Training method where weight mask is updated randomly periodically

Similar Papers 제목 키워드 기반

On the Scaling of Robustness and Effectiveness in Dense Retrieval

2025-05-30 · Yu-An Liu, Ruqing Zhang, Jiafeng Guo, Maarten de Rijke 외

Robustness and Effectiveness are critical aspects of developing dense retrieval models for real-world applications. It is known that there is a trade-off between the two. Recent work has addressed scaling laws of effecti…

Adversarial RobustnessRetrieval

Scaling Laws for Embedding Dimension in Information Retrieval

2026-02-04 · Julian Killingback, Mahta Rafiee, Madine Manas, Hamed Zamani arxiv

Dense retrieval, which encodes queries and documents into a single dense vector, has become the dominant neural retrieval approach due to its simplicity and compatibility with fast approximate nearest neighbor algorithms…

Information Retrieval

Generalizing Scaling Laws for Dense and Sparse Large Language Models

2025-08-08 · Md Arafat Hossain, Xingfu Wu, Valerie Taylor, Ali Jannesari arxiv

Despite recent advancements of large language models (LLMs), optimally predicting the model size for LLM pretraining or allocating optimal resources still remains a challenge. Several efforts have addressed the challenge…

Scaling Laws for Robust Comparison of Open Foundation Language-Vision Models and Datasets

2025-06-05 · Marianna Nezhurina, Tomer Porian, Giovanni Pucceti, Tommie Kerssies 외

In studies of transferable learning, scaling laws are obtained for various important foundation models to predict their properties and performance at larger scales. We show here how scaling law derivation can also be use…

Scaling Laws for Online Advertisement Retrieval

2024-11-20 · Yunli Wang, Zixuan Yang, Zhen Zhang, Zhiqiang Wang 외

The scaling law is a notable property of neural network models and has significantly propelled the development of large language models. Scaling laws hold great promise in guiding model design and resource allocation. Re…

Retrieval