paper-with-me

홈 › Papers

Inside the Scaffold: A Source-Code Taxonomy of Coding Agent Architectures

2026-04-03 · Benjamin Rombaut arxiv

LLM-based coding agents can localize bugs, generate patches, and run tests with diminishing human oversight, yet the scaffolding code that surrounds the language model (the control loop, tool definitions, state management, and context strategy) remains poorly understood. Existing surveys classify agents by abstract capabilities (tool use, planning, reflection) that cannot distinguish between architecturally distinct systems, and trajectory studies observe what agents do without examining the scaffold code that determines why. This paper presents a source-code-level architectural taxonomy derived from analysis of 13 open-source coding agent scaffolds at pinned commit hashes. Each agent is characterized across 12 dimensions organized into three layers: control architecture, tool and environment interface, and resource management. The analysis reveals that scaffold architectures resist discrete classification: control strategies range from fixed pipelines to Monte Carlo Tree Search, tool counts range from 0 to 37, and context compaction spans seven distinct strategies. Five loop primitives (ReAct, generate-test-repair, plan-execute, multi-attempt retry, tree search) function as composable building blocks that agents layer in different combinations; 11 of 13 agents compose multiple primitives rather than relying on a single control structure. Dimensions converge where external constraints dominate (tool capability categories, edit formats, execution isolation) and diverge where open design questions remain (context compaction, state management, multi-model routing). All taxonomic claims are grounded in file paths and line numbers, providing a reusable reference for researchers studying agent behavior and practitioners designing new scaffolds.

📄 PDF Abstract BibTeX arXiv:2604.03515

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Beyond Surface Forms: A Comprehensive, Mechanism-Oriented Taxonomy of Indirect Linguistic Encoding for LLM-Based Coded Language Detection

2026-06-25 · Hamid Reza Firoozfar, Mohammadsadegh Abolhasani, Reza Mousavi, Paul Jen-Hwa Hu arxiv

To avoid moderation and surveillance on social media, some users routinely invent indirect linguistic expressions (ILE) that camouflage sensitive meanings. Such expressions surface as algospeak, euphemisms, and adversari…

Sandboxed Coding Agents are Competitive Omni-modal Task Solvers

2026-05-30 · Dongping Chen, Xuanao Huang, Zhihan Hu, Qingyuan Shi 외 arxiv

As multimodal LLMs increasingly target video and audio, it is often assumed that such tasks require native omnimodal models. We show that this is not always the case: coding agents with only text+image access and a sandb…

Don't Blame the Large Language Model: How Scaffolding Evolution Shapes Coding Agent Quality

2026-07-04 · Oussama Ben Sghaier, Hao Li, Bram Adams, Ahmed E. Hassan arxiv

Coding agents, autonomous systems that use large language models (LLMs) to resolve software engineering tasks, rely on agentic scaffolding: a middleware layer in between a developer and a large language model that orches…

Promoting Graph Awareness in Linearized Graph-to-Text Generation

2020-12-31 · Findings (ACL) 2021 8 · Alexander Hoyle, Ana Marasović, Noah Smith

Generating text from structured inputs, such as meaning representations or RDF triples, has often involved the use of specialized graph-encoding neural networks. However, recent applications of pretrained transformers to…

DenoisingText Generation

Model or Harness? An Interaction-Centric Taxonomy for Localizing Agent Failures

2026-07-30 · Harsh Raj, Vipul Gupta, Anas Mahmoud, Razvan-Gabriel Dumitru 외 hf

Existing evaluations often reduce agent failures to system-level outcomes, obscuring where the fault originated and which intervention would improve the agent system. This creates a repair-assignment problem: the same vi…