paper-with-me

홈 › Papers

Practical Code RAG at Scale: Task-Aware Retrieval Design Choices under Compute Budgets

2025-10-23 · Timur Galimzyanov, Olga Kolomyttseva, Egor Bogomolov arxiv

We study retrieval design for code-focused generation tasks under realistic compute budgets. Using two complementary tasks from Long Code Arena -- code completion and bug localization -- we systematically compare retrieval configurations across various context window sizes along three axes: (i) chunking strategy, (ii) similarity scoring, and (iii) splitting granularity. (1) For PL-PL, sparse BM25 with word-level splitting is the most effective and practical, significantly outperforming dense alternatives while being an order of magnitude faster. (2) For NL-PL, proprietary dense encoders (Voyager-3 family) consistently beat sparse retrievers, however requiring 100x larger latency. (3) Optimal chunk size scales with available context: 32-64 line chunks work best at small budgets, and whole-file retrieval becomes competitive at 16000 tokens. (4) Simple line-based chunking matches syntax-aware splitting across budgets. (5) Retrieval latency varies by up to 200x across configurations; BPE-based splitting is needlessly slow, and BM25 + word splitting offers the best quality-latency trade-off. Thus, we provide evidence-based recommendations for implementing effective code-oriented RAG systems based on task requirements, model constraints, and computational efficiency.

📄 PDF Abstract BibTeX arXiv:2510.20609

Code (0)

등록된 구현이 없습니다.

Tasks

Computational EfficiencyCode Completion

Similar Papers 제목 키워드 기반

Attribute-Aware Deep Hashing with Self-Consistency for Large-Scale Fine-Grained Image Retrieval

2023-11-21 · Xiu-Shen Wei, Yang shen, Xuhao Sun, Peng Wang 외

Our work focuses on tackling large-scale fine-grained image retrieval as ranking the images depicting the concept of interests (i.e., the same sub-category labels) highest based on the fine-grained details in the query. …

AttributeDeep HashingImage ReconstructionImage Retrieval+1

A$^2$-Net: Learning Attribute-Aware Hash Codes for Large-Scale Fine-Grained Image Retrieval

2021-12-01 · NeurIPS 2021 12 · Xiu-Shen Wei, Yang shen, Xuhao Sun, Han-Jia Ye 외

Our work focuses on tackling large-scale fine-grained image retrieval as ranking the images depicting the concept of interests (i.e., the same sub-category labels) highest based on the fine-grained details in the query. …

AttributeDecoderImage RetrievalRetrieval

Do We Need Bigger Models for Science? Task-Aware Retrieval with Small Language Models

2026-04-02 · Florian Kelber, Matthias Jobst, Yuni Susanti, Michael Färber arxiv

Scientific knowledge discovery increasingly relies on large language models, yet many existing scholarly assistants depend on proprietary systems with tens or hundreds of billions of parameters. Such reliance limits repr…

Question Answering

cAST: Enhancing Code Retrieval-Augmented Generation with Structural Chunking via Abstract Syntax Tree

2025-06-18 · Yilin Zhang, Xinran Zhao, Zora Zhiruo Wang, Chenyang Yang 외

Retrieval-Augmented Generation (RAG) has become essential for large-scale code generation, grounding predictions in external code corpora to improve actuality. However, a critical yet underexplored aspect of RAG pipeline…

ChunkingCode GenerationRAGRetrieval+1

CiCo: Domain-Aware Sign Language Retrieval via Cross-Lingual Contrastive Learning

2023-03-22 · CVPR 2023 1 · Yiting Cheng, Fangyun Wei, Jianmin Bao, Dong Chen 외

This work focuses on sign language retrieval-a recently proposed task for sign language understanding. Sign language retrieval consists of two sub-tasks: text-to-sign-video (T2V) retrieval and sign-video-to-text (V2T) re…

Contrastive LearningRetrievalSign Language Retrievalspeech-recognition+3