paper-with-me

Papers

CodeTeam: An LLM-Powered Multi-Agent Framework for Repository-Level Code Generation

2026-06-20 · Yifei Wang, Ruiyin Li, Peng Liang, Qiong Feng, Zengyang Li, Mojtaba Shahin, Arif Ali Khan arxiv

Natural language to repository generation (NL2Repo) requires a system to construct an entire software repository from a natural-language requirements document. Compared with function-level code generation, this task demands longer planning horizons, stable interfaces across files, and iterative debugging of cross-file inconsistencies. To address these challenges, we propose CodeTeam, an LLM-based multi-agent framework that separates planning, decision making, and implementation into distinct, coordinated stages. In the planning stage, multiple Architect agents draft competing software design sketches (SDS), optionally grounded by retrieved design references. A CTO agent then evaluates, selects, and normalizes the most promising SDS into a machine-checkable contract that specifies file ownership, public interfaces, and dependency constraints. In the implementation stage, Developer agents generate code under a dependency-aware scheduler with bounded context and lightweight Git-based coordination, while a QA agent runs tests and drives iterative repairs. On the synthesis-based SketchEval benchmark, we explicitly compare CodeTeam's prompt-engineering (PE) and supervised fine-tuning (SFT) variants with the corresponding CodeS variants, where CodeTeam improves the overall SketchBLEU by 4.1 and 2.9 absolute points, respectively. On the execution-based NL2Repo-Bench benchmark, used as an external validation protocol, CodeTeam achieves the highest average test pass rate in both settings (34.6% PE, 42.3% SFT), confirming that the sketch-improvements extend to functional correctness under upstream test suites. Ablation results show that project-specific developer allocation and retrieval-augmented planning each contribute substantially to the SketchBLEU improvement (9.9% and 8.1% relative, respectively). CodeTeam and the experimental results are available at https://github.com/WhitenWhiten/CodeTeam

📄 PDF Abstract BibTeX arXiv:2606.22082

Code (0)

등록된 구현이 없습니다.

Tasks

Code GenerationDecision Making

Similar Papers 제목 키워드 기반

RepoAgent: An LLM-Powered Open-Source Framework for Repository-level Code Documentation Generation

2024-02-26 · Qinyu Luo, Yining Ye, Shihao Liang, Zhong Zhang 외

Generative models have demonstrated considerable potential in software engineering, particularly in tasks such as code generation and debugging. However, their utilization in the domain of code documentation generation r…

Code Documentation GenerationCode GenerationLanguage ModelingLanguage Modelling+1

RepoAtlas: Guiding Coding Agents via Evolving Multimodal Repository Views

2026-09-15 · Yunxiang Zhang, Haiquan Wang, JiaWei Guo, Hanyang Xia 외 arxiv

Large language model (LLM)-powered coding agents have made rapid progress in automating software engineering tasks, yet repository-level issue resolution remains challenging. Beyond generating a plausible patch, an agent…

ArchAgent: Scalable Legacy Software Architecture Recovery with LLMs

2026-01-19 · Rusheng Pan, Bingcheng Mao, Tianyi Ma, Zhenhua Ling arxiv

Recovering accurate architecture from large-scale legacy software is hindered by architectural drift, missing relations, and the limited context of Large Language Models (LLMs). We present ArchAgent, a scalable agent-bas…

SiriuS: Self-improving Multi-agent Systems via Bootstrapped Reasoning

2025-02-07 · Wanjia Zhao, Mert Yuksekgonul, Shirley Wu, James Zou

Multi-agent AI systems powered by large language models (LLMs) are increasingly applied to solve complex tasks. However, these systems often rely on fragile, manually designed prompts and heuristics, making optimization …

DrunkAgent: Stealthy Memory Corruption in LLM-Powered Recommender Agents

2025-03-31 · Shiyi Yang, Zhibo Hu, Xinshu Li, Chen Wang 외

Large language model (LLM)-powered agents are increasingly used in recommender systems (RSs) to achieve personalized behavior modeling, where the memory mechanism plays a pivotal role in enabling the agents to autonomous…

Collaborative FilteringLarge Language ModelRecommendation Systems