paper-with-me

Code Search

7개 벤치마크 · 논문 147편 · 이 태스크의 논문 보기 →

Benchmarks

CodeSearchNet

결과 6개

CoDesc

결과 3개

CodeXGLUE - AdvTest

결과 3개

결과 1개

CoIR

결과 1개

CodeSearchNet - Ruby

결과 1개

Most implemented

Papers

MediaWiki Code2Code Search: Neural Retrieval for the Semantic Discovery of Open-Source Software Entities

2026-07-29 · Francesco Tosoni arxiv

Code search in large-scale ecosystems is often hindered by the lexical gap between user queries and implementation details, alongside the trade-off between the low latency of traditional Information Retrieval (IR) and th…

Information RetrievalCode Search

Harness Handbook: Making Evolving Agent Harnesses Readable,Navigable, and Editable

2026-07-14 · Ruhan Wang, Yucheng Shi, Zongxia Li, Zhongzhi Li 외 hf

The capability of a modern AI agent depends not only on its foundation model but also on its harness, which constructs prompts, manages state, invokes tools, and coordinates execution. As models, APIs, environments, and …

Code Search

Scientific Code Search at Scale: A Multi-Domain Dataset and Benchmark

2026-07-03 · Nishan Pantha, Pranath Reddy Kumbam, Sajil Awale, Pushwitha Krishnappa 외 arxiv

Scientists increasingly rely on open-source tools to support their research workflows, yet discovering relevant software among over 600 million GitHub repositories remains challenging. Existing code search benchmarks foc…

Information RetrievalCode Search

Decoupling Code Complexity from Newcomer Participation: A Causal Study of AI Coding Agent Adoption in OSS

2026-07-02 · Weiwei Xu, Xuanning Cui, Hengzhi Ye, Minghui Zhou arxiv

Open-source projects depend on a steady inflow of newcomers. A growing concern is that AI coding agents (tools such as Cursor and Claude Code that write code from natural-language instructions) will crowd them out, by ab…

Code Search

UniCoder: Unified Visual-to-Code Generation via Symbolic Rewards and Reference-Guided Code Optimization

2026-06-30 · Yaozhi Zheng, Yilei Jiang, Manyuan Zhang, Yuxuan Wan 외 arxiv

Visual-to-Code generation, which transforms scientific plots, vector graphics, and webpages into executable scripts, demands a level of pixel-precise alignment that standard Multimodal Large Language Models (MLLMs) fail …

Reinforcement LearningCode GenerationCode Search

Recall Before Rerank: Benchmarking Deep Learning Models for Large-Scale Code-to-Code Retrieval

2026-06-24 · Leonardo Venuta, Francesco Tosoni, Paolo Ferragina arxiv

Semantic code search and clone detection are essential for software development, maintenance, and reuse. This paper evaluates the effectiveness, efficiency, and scalability of contemporary deep learning models for first-…

Code Search

전체 147편 보기 →