paper-with-me

홈 › Papers

Agent Primitives: Reusable Latent Building Blocks for Multi-Agent Systems

2026-02-03 · Haibo Jin, Peng Kuang, Ye Yu, Xiaopeng Yuan, Haohan Wang arxiv

While existing multi-agent systems (MAS) can handle complex problems by enabling collaboration among multiple agents, they are often highly task-specific, relying on manually crafted agent roles and interaction prompts, which leads to increased architectural complexity and limited reusability across tasks. Moreover, most MAS communicate primarily through natural language, making them vulnerable to error accumulation and instability in long-context, multi-stage interactions within internal agent histories. In this work, we propose \textbf{Agent Primitives}, a set of reusable latent building blocks for LLM-based MAS. Inspired by neural network design, where complex models are built from reusable components, we observe that many existing MAS architectures can be decomposed into a small number of recurring internal computation patterns. Based on this observation, we instantiate three primitives: Review, Voting and Selection, and Planning and Execution. All primitives communicate internally via key-value (KV) cache, which improves both robustness and efficiency by mitigating information degradation across multi-stage interactions. To enable automatic system construction, an Organizer agent selects and composes primitives for each query, guided by a lightweight knowledge pool of previously successful configurations, forming a primitive-based MAS. Experiments show that primitives-based MAS improve average accuracy by 12.0-16.5\% over single-agent baselines, reduce token usage and inference latency by approximately 3$\times$-4$\times$ compared to text-based MAS, while incurring only 1.3$\times$-1.6$\times$ overhead relative to single-agent inference and providing more stable performance across model backbones.

📄 PDF Abstract BibTeX arXiv:2602.03695

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Algorithmic Primitives and Compositional Geometry of Reasoning in Language Models

2025-10-13 · Samuel Lippl, Thomas McGee, Kimberly Lopez, Ziwen Pan 외 arxiv

How do latent and inference time computations enable large language models (LLMs) to solve multi-step reasoning? We introduce a framework for tracing and steering algorithmic primitives that underlie model reasoning. Our…

Latent Actions from Factorized Transition Effects under Agent Ambiguity

2026-06-29 · Heejeong Nam, Chandradithya S Jonnalagadda, Harshit Aggarwal, Eric Xu 외 arxiv

Latent Action Models (LAMs) learn action-like proxies from observation transitions. However, in multi-object or distractor-rich scenes, these visual effects mix agent motion with distractors, camera dynamics, and backgro…

Composing Verifiable Conceptual Models via Building Blocks: Towards Design-Time Verification of Agentic AI Workflows

2026-06-19 · Noe Y. Flandre, Alexander C. Nwala, Philippe J. Giabbanelli arxiv

Agentic AI systems orchestrate multiple LLM-based agents through workflow architectures that coordinate decisions, tools, and external actions. While current platforms emphasize runtime safeguards, little support exists …

cuSLINK: Single-linkage Agglomerative Clustering on the GPU

2023-06-28 · Corey J. Nolet, Divye Gala, Alex Fender, Mahesh Doijade 외

In this paper, we propose cuSLINK, a novel and state-of-the-art reformulation of the SLINK algorithm on the GPU which requires only $O(Nk)$ space and uses a parameter $k$ to trade off space and time. We also propose a se…

ClusteringGPUgraph construction

Inside the Scaffold: A Source-Code Taxonomy of Coding Agent Architectures

2026-04-03 · Benjamin Rombaut arxiv

LLM-based coding agents can localize bugs, generate patches, and run tests with diminishing human oversight, yet the scaffolding code that surrounds the language model (the control loop, tool definitions, state managemen…